arXiv · 2502.03249
Portable Lattice QCD implementation based on OpenCL
Abstract
The presence of GPU from different vendors demands the Lattice QCD codes to support multiple architectures. To this end, Open Computing Language (OpenCL) is one of the viable frameworks for writing a portable code. It is of interest to find out how the OpenCL implementation performs as compared to the code based on a dedicated programming interface such as CUDA for Nvidia GPUs. We have developed an OpenCL backend for our already existing code of the Wuppertal-Budapest collaboration. In this contribution, we show benchmarks of the most time consuming part of the numerical simulation, namely, the inversion of the Dirac operator. We present the code performance on the JUWELS and LUMI Supercomputers based on Nvidia and AMD graphics cards, respectively, and compare with the CUDA backend implementation.
Explore related subjects
Keep this discovery
Piyush Kumar, Szabolcs Borsanyi, Jana N. Guenther, Chik Him Wong. 2025-02-05. Portable Lattice QCD implementation based on OpenCL. https://arxiv.org/abs/2502.03249
Cite the original work for its findings. Save a collection to share your selection of sources.