Improving Processor and Cache Locality in Fine-Grain Parallel Computations using Object-Affinity Scheduling and Continuation Passing
Robert J. Fowler, Leonidas I. Kontothanassis · 1992
olychronopoulos. Auto-scheduling: Control flow and data flow come together. Technical Report CSRD RPT 1058, Center for Supercomputing Research and Development, University of Illinois at Urbana-Champaign, December 1990. M. S. Squillante. Issues in Shared-Memory Multiprocessor Scheduling: A Performance Evaluation. PhD thesis, Department of Computer Science and Engineering, University of Washington, October 1990. R. H. Thomas and W. Crowther. The Uniform System: An approach to runtime support for large scale shared memory parallel processors. In Proceedings of the 1988 International Conference on Parallel Processing, pages 245-254, August 1988. T. H. Tzen and L. M. Ni. Trapezoid self-scheduling: A practical scheduling scheme for parallel compilers. Technical Report MSU-CPS-ACS-27, Michigan State University, September 1991. L. G. Valiant. A bridging model for parallel computation. Communications of the A CM, 33(8), August 1990. Raj Vaswani and John Zahorjan. The implications of cache