Exploiting Vector Processing in Dynamic Binary Translation
Chih‐Min Lin, Sheng‐Yu Fu, Ding‐Yong Hong, Yuping Liu, Jan‐Jan Wu, Wei‐Chung Hsu · 2019
Auto vectorization techniques have been adopted by compilers to exploit data-level parallelism in parallel processing for decades. However, since processor architectures have kept enhancing with new features to improve vector/SIMD performance, legacy application binaries failed to fully exploit new vector/SIMD capabilities in modern architectures. For example, legacy ARMv7 binaries cannot benefit from ARMv8 SIMD double precision capability, and legacy x86 binaries cannot enjoy the power of AVX-512 extensions.