arrow
返回

Exploiting Memory Access Patterns to Improve Memory Performance in Data-Parallel Architectures

delete2011-01-01
delete141
PRE
AI
B
Byunghyun Jang *
D
David Kaeli
DOI:10.1109/TPDS.2010.107delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
The introduction of General-Purpose computation on GPUs (GPGPUs) has changed the landscape for the future of parallel computing. At the core of this phenomenon are massively multithreaded, data-parallel architectures possessing impressive acceleration ratings, offering low-cost supercomputing together with attractive power budgets. Even given the numerous benefits provided by GPGPUs, there remain a number of barriers that delay wider adoption of these architectures. One major issue is the heterogeneous and distributed nature of the memory subsystem commonly found on data-parallel architectures. Application acceleration is highly dependent on being able to utilize the memory subsystem effectively so that all execution units remain busy. In this paper, we present techniques for enhancing the memory efficiency of applications on data-parallel architectures, based on the analysis and characterization of memory access patterns in loop bodies; we target vectorization via data transformation to benefit vector-based architectures (e. g., AMD GPUs) and algorithmic memory selection for scalar-based architectures (e. g., NVIDIA GPUs). We demonstrate the effectiveness of our proposed methods with kernels from a wide range of benchmark suites. For the benchmark kernels studied, we achieve consistent and significant performance improvements (up to 11.4 x and 13.5 x over baseline GPU implementations on each platform, respectively) by applying our proposed methodology.
Keyword:
General-purpose computation on GPUs (GPGPUs)
GPU computing
memory optimization
memory access pattern
vectorization
memory selection
memory coalescing
data parallelism
data-parallel architectures
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

IEEE Transactions on Parallel and Distributed Systems 封面图
IEEE Transactions on Parallel and Distributed Systems
IF:
6
论文数:
5.2K
被引数:
1.1W

机构

N
Northeastern University
学者数:
2.5W
论文数: 1.6W
被引数: 3.0W
引用论文

引用论文

err分享
err收藏
PV-specific loss of the transcriptional coactivator PGC-1α slows down the evolution of epileptic activity in an acute ictogenic model
err2022-01-01
err0
errOAAI
errConnie Mackenzie-Gray Scott; R. Ryley Parrish; Darren Walsh; Claudia Racca; Rita M. Cowell; Andrew J. Trevelyan
err分享
err收藏
Anomalous Magnetoresistance in the π-d System (DIETSe)2FeCl4
err2007-01-31
err0
PREAI
errMitsuhiko Maesato; Tomohito Kawashima; Gunzi Saito; Takashi Shirahata; Megumi Kibune; Tatsuro Imakubo
err分享
err收藏
err2006-06-01
err0
PREAI
errR. De Giorgio
err分享
err收藏
Cognitive Functioning in Coronary Artery Disease Patients: Associations with Thyroid Hormones, N-Terminal Pro-B-Type Natriuretic Peptide and High-Sensitivity C-Reactive Protein
err2017-01-23
err0
errOAAI
errJulius Burkauskas; Adomas Bunevicius; Julija Brozaitiene; Julius Neverauskas; Peter Lang; Robert Duwors; Narseta Mickuviene; Robertas Bunevicius
err分享
err收藏
学者 查看更多内容