Inference of ML models on Intel GPUs with SYCL and Intel OneAPI using SOFIE

Ioanna-Maria Panagou, L. Moneta, Sanjiban Sengupta · Zenodo (CERN European Organization for Nuclear Research) · 2023

TMVA provides a fast inference system that takes an ONNX model as input and produces compilationready standalone C++ scripts as output which can be compiled and executed on CPU architectures. The idea of this project is to extend this capability to generate from the TMVA SOFIE model representation code that can be run also on Intel GPUs using both SYCL and Intel OneAPI libraries. These will allow for a more efficient evaluation of these models on Intel accelerator hardware.

Read the paper · More papers on PaperTik