An Exploration of OpenCL for a Numerical Relativity Application
arXiv:1010.3816
Abstract
Currently there is considerable interest in making use of many-core processor architectures, such as Nvidia and AMD graphics processing units (GPUs) for scientific computing. In this work we explore the use of the Open Computing Language (OpenCL) for a typical Numerical Relativity application: a time-domain Teukolsky equation solver (a linear, hyperbolic, partial differential equation solver using finite-differencing). OpenCL is the only vendor-agnostic and multi-platform parallel computing framework that has been adopted by all major processor vendors. Therefore, it allows us to write portable source-code and run it on a wide variety of compute hardware and perform meaningful comparisons. The outcome of our experimentation suggests that it is relatively straightforward to obtain order-of-magnitude gains in overall application performance by making use of many-core GPUs over multi-core CPUs and this fact is largely independent of the specific hardware architecture and vendor. We also observe that a single high-end GPU can match the performance of a small-sized, message-passing based CPU cluster.
6 pages, 1 table; Accepted for publication in Parallel and Distributed Computing and Systems (PDCS 2011)
References in corpus (8)
- A Performance Comparison of CUDA and OpenCL
- Universality of massive scalar field late-time tails in black-hole spacetimes
- Massive scalar field instability in Kerr spacetime
- Late-time Kerr tails revisited
- Numerical modeling of gravitational wave sources accelerated by OpenCL
- Collision of spinning black holes in the close limit
- High-Precision Numerical Simulations of Rotating Black Holes Accelerated by CUDA
- An exploration of CUDA and CBEA for a gravitational wave source-modelling application