Tales of the Tail
Jialin Li, Naveen Sharma, Dan R. K. Ports, Steven D. Gribble · 2014
Interactive services often have large-scale parallel implementations. To deliver fast responses, the median and tail latencies of a service's components must be low. In this paper, we explore the hardware, OS, and application-level sources of poor tail latency in high throughput servers executing on multi-core machines.