Performance optimizations for porting the openQD package to GPUs
arXiv:2202.07388
Abstract
OpenQD code has been used by the RC collaboration for the generation of fully dynamical QCD+QED gauge configurations with C boundary conditions. In this talk, optimization of solvers provided with the openQD package relevant for porting the code on GPU-accelerated supercomputing platforms is discussed. We present the analysis of the current implementations of the GCR solver preconditioned with Schwarz alternating procedure for ill-conditioned Dirac-operators. With the goal of enabling support for GPUs from various vendors, a novel method of adaptive CPU/GPU-hybrid implementation is proposed.