Enabling Profiling For SYCL Applications

Callum Fare · 2018

Since GPGPU devices have become mainstream, more and more software is being written to target many-core devices. Developers are now required to think in parallel in order to run applications with maximum performance, however, the ability to target a wide range of devices is vital. GPUs can range from very powerful discrete cards to extremely low-power embedded chips and, to be efficient, developers must be able to reuse their code in different scenarios. OpenCL™ addresses this issue by providing a C like programming language that can target different architectures, however, it requires a deep knowledge of the underlying hardware to be used efficiently. SYCL™ provides a C++ abstraction layer that simplifies parallel development, allowing developers to leverage the power of OpenCL™ while reducing the amount of code required.

Read the paper · More papers on PaperTik