High throughput heavy hitter aggregation for modern SIMD processors
Orestis Polychroniou, Kenneth Andrew Ross · 2013
Heavy hitters are data items that occur at high frequency in a data set. They are among the most important items for an organization to summarize and understand during analytical processing. In data sets with sufficient skew, the number of heavy hitters can be relatively small. We take advantage of this small footprint to compute aggregate functions for the heavy hitters in fast cache memory in a single pass.