Lightweight Deep Learning Applications on AVX-512
André Ramos Carneiro, Matheus S. Serpa, Philippe O. A. Navaux · 2021 IEEE Symposium on Computers and Communications (ISCC) · 2021
Machine Learning and Deep Learning applications are of paramount importance these days. Different areas of academia and industry use daily workloads based on these applications. Several aspects are relevant regarding their applicability, such as the complexity and accuracy of the models and their performance and energy efficiency. Currently, there is a trend to usually favor the use of GPUs to train and execute Deep Learning models, intensified by specialized hardware. However, this article demonstrates that using a CPU with AVX-512 instructions can achieve comparable performance to current GPUs and, depending on the workload, suppress it by ≈ 1.8x.