Optimizing MapReduce for GPUs with effective shared memory usage
Linchuan Chen, Gagan Agrawal · 2012
Accelerators and heterogeneous architectures in general, and GPUs in particular, have recently emerged as major players in high performance computing. For many classes of applications, MapReduce has emerged as the framework for easing parallel programming and improving programmer productivity. There have already been several efforts on implementing MapReduce on GPUs.