Return
Domain-Specific Optimization and Generation of High-Performance GPU Code for Stencil Computations
DOI:10.1109/JPROC.2018.2862896.png)
Abstract
En 中文
Stencil computations arise in a number of computational domains. They exhibit significant data parallelism and are thus well suited for execution on graphical processing units (GPUs), but can be memory-bandwidth limited unless temporal locality is utilized via tiling. This paper describes how effective tiled code can be generated for GPUs from a domain-specific language (DSL) for stencils. Experimental results demonstrate the benefits of such a domain-specific optimization approach over state-of-the-art general-purpose compiler optimizations.
Keywords:
General-purpose graphical processing unit (GPGPU)
resource optimization
stencil computations
streaming
tiling
AI Summary
Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.
Journal
IF:
25.9
Papers:
9.9K
Citations:
4.5W

