Design and Performance of an Intel Xeon Phi based Cluster for Reverse Time Migration

V. Arslan, J.Y. Blanc, Marc Tchiboukdjian, Philippe Thierry, G. Thomas-Collignon · Proceedings · 2014

In this paper, we design an Intel Xeon Phi based cluster specifically tuned for RTM and compare its performance to our current optimized architecture consisting of Nvidia GPU accelerated nodes. The Xeon Phi nodes are designed to offer a good balance between co-processor computing capabilities, host computing capabilities, local scratch bandwidth and network bandwidth. Performance is evaluated at the system level including wave propagation kernels but also wavefield checkpointing to the local scratch and pre and post-processing steps. Moreover, the kernels are tuned for the Xeon Phi and compared to our GPU optimized kernels taking into account the tuning effort. Overall, the obtained performance is comparable to our GPU-based architecture while offering a better portability since most of the code is identical to the CPU implementation.

Read the paper · More papers on PaperTik