Return
Mesh-particle interpolations on graphics processing units and multicore central processing units
DOI:10.1098/rsta.2011.0074.png)
Abstract
En 中文
Particle-mesh interpolations are fundamental operations for particle-in-cell codes, as implemented in vortex methods, plasma dynamics and electrostatics simulations. In these simulations, the mesh is used to solve the field equations and the gradients of the fields are used in order to advance the particles. The time integration of particle trajectories is performed through an extensive resampling of the flow field at the particle locations. The computational performance of this resampling turns out to be limited by the memory bandwidth of the underlying computer architecture. We investigate how mesh-particle interpolation can be efficiently performed on graphics processing units (GPUs) and multicore central processing units (CPUs), and we present two implementation techniques. The single-precision results for the multicore CPU implementation show an acceleration of 45-70x, depending on system size, and an acceleration of 85-155x for the GPU implementation over an efficient single-threaded C++ implementation. In double precision, we observe a performance improvement of 30-40x for the multicore CPU implementation and 20-45x for the GPU implementation. With respect to the 16-threaded standard C++ implementation, the present CPU technique leads to a performance increase of roughly 2.8-3.7x in single precision and 1.7-2.4x in double precision, whereas the GPU technique leads to an improvement of 9x in single precision and 2.2-2.8x in double precision.
Keywords:
central processing units
graphics processing units
high-performance computing
mesh-particle
AI Summary
Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.
Journal
P
IF:
3.7
Papers:
7.7K
Citations:
2.8W

