High throughput heavy hitter aggregation for modern SIMD processors

Orestis Polychroniou, Kenneth Andrew Ross · 2013

Heavy hitters are data items that occur at high frequency in a data set. They are among the most important items for an organization to summarize and understand during analytical processing. In data sets with sufficient skew, the number of heavy hitters can be relatively small. We take advantage of this small footprint to compute aggregate functions for the heavy hitters in fast cache memory in a single pass.

Read the paper · More papers on PaperTik