General purpose lattice QCD code set Bridge++ 2.0 for high performance computing
arXiv:2111.04457 · doi:10.1088/1742-6596/2207/1/012053
Abstract
Bridge++ is a general-purpose code set for a numerical simulation of lattice QCD aiming at a readable, extensible, and portable code while keeping practically high performance. The previous version of Bridge++ is implemented in double precision with a fixed data layout. To exploit the high arithmetic capability of new processor architecture, we extend the Bridge++ code so that optimized code is available as a new branch, i.e., an alternative to the original code. This paper explains our strategy of implementation and displays application examples to the following architectures and systems: Intel AVX-512 on Xeon Phi Knights Landing, Arm A64FX-SVE on Fujitsu A64FX (Fugaku), NEC SX-Aurora TSUBASA, and GPU cluster with NVIDIA V100.
6 pages, 6 figures. Talk by I.Kanamori at the XXXII IUPAP Conference on Computational Physics (CCP2021), Coventry, UK, 1-5 August 2021
References in corpus (1)
Cited by in corpus (5)
- Bridge++ 2.0: Benchmark results on supercomputer Fugaku
- Wilson matrix kernel for lattice QCD on A64FX architecture
- Mixed precision solvers with half-precision floating point numbers for Lattice QCD on A64FX processor
- Accelerating iterative linear equation solver using modified domain-wall fermion matrix in lattice QCD simulations
- The dynamics of zero modes in lattice gauge theory -- difference between SU(2) and SU(3) in 4D