Compute Unified Device Architecture

Wen‐mei Hwu, David B. Kirk · The MIT Press eBooks · 2015

This chapter contains sections titled: 15.1 A Brief History Leading to CUDA, 15.2 CUDA Program Structure, 15.3 A Vector Addition Example, 15.4 Device Memories and Data Transfer, 15.5 Kernel Functions and Threading, 15.6 More on CUDA Thread Organization, 15.7 Mapping Threads to Multidimensional Data, 15.8 Synchronization and Transparent Scalability, 15.9 Assigning Resources to Blocks, 15.10 CUDA Streams and Task Parallelism, 15.11 Summary

Read the paper · More papers on PaperTik