Many-gluon tree amplitudes on modern GPUs: A case study for novel event generators
arXiv:2106.06507
Abstract
The compute efficiency of Monte-Carlo event generators for the Large Hadron Collider is expected to become a major bottleneck for simulations in the high-luminosity phase. Aiming at the development of a full-fledged generator for modern GPUs, we study the performance of various recursive strategies to compute multi-gluon tree-level amplitudes. We investigate the scaling of the algorithms on both CPU and GPU hardware. Finally, we provide practical recommendations as well as baseline implementations for the development of future simulation programs. The GPU implementations can be found at: https://www.gitlab.com/ebothmann/blockgen-archive.
27 pages, 9 figures, revised version, Submission to SciPost
References in corpus (7)
- The automated computation of tree-level and next-to-leading order differential cross sections, and their matching to parton shower simulations
- An Automated Implementation of On-Shell Methods for One-Loop Amplitudes
- GoSam-2.0: a tool for automated one-loop calculations within the Standard Model and beyond
- Monte Carlo integration on GPU
- Accelerated Matrix Element Method with Parallel Computing
- Thread-Scalable Evaluation of Multi-Jet Observables
- Comparing efficient computation methods for massless QCD tree amplitudes: Closed Analytic Formulae versus Berends-Giele Recursion