arrow
返回

Efficient GPU Spatial-Temporal Multitasking

delete2015-03-01
delete79
PRE
AI
Y
Yun Liang *
K
Kyle Rupnow
R
Rick Siow Mong Goh
D
Deming Chen
DOI:10.1109/TPDS.2014.2313342delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Heterogeneous computing nodes are now pervasive throughout computing, and GPUs have emerged as a leading computing device for application acceleration. GPUs have tremendous computing potential for data-parallel applications, and the emergence of GPUs has led to proliferation of GPU-accelerated applications. This proliferation has also led to systems in which many applications are competing for access to GPU resources, and efficient utilization of the GPU resources is critical to system performance. Prior techniques of temporal multitasking can be employed with GPU resources as well, but not all GPU kernels make full use of the GPU resources. There is, therefore, an unmet need for spatial multitasking in GPUs. Resources used inefficiently by one kernel can be instead assigned to another kernel that can more effectively use the resources. In this paper we propose a software-hardware solution for efficient spatial-temporal multitasking and a software based emulation framework for our system. We pair an efficient heuristic in software with hardware leaky-bucket based thread-block interleaving to implement spatial-temporal multitasking. We demonstrate our techniques on various GPU architecture using nine representative benchmarks from CUDA SDK. Our experiments on Fermi GTX480 demonstrate performance improvement by up to 46% (average 26%) over sequential GPU task execution and 37% (average 18%) over default concurrent multitasking. Compared with the state-of-the-art Kepler K20 using Hyper-Q technology, our technique achieves up to 40% (average 17%) performance improvement over default concurrent multitasking.
Keyword:
GPU
spatial
temporal
multitasking
resource allocation
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

IEEE Transactions on Parallel and Distributed Systems 封面图
IEEE Transactions on Parallel and Distributed Systems
IF:
6
论文数:
5.2K
被引数:
1.1W

机构

A
a*star - institute of high performance computing (ihpc)
学者数:
1.5K
论文数: 1.3K
被引数: 3
University of Illinois System 封面图
University of Illinois System
学者数:
6.8W
论文数: 6.2W
被引数: 644
P
peking university
学者数:
11.9W
论文数: 8.7W
被引数: 146
A
agency for science technology & research (a*star)
学者数:
2.2W
论文数: 1.9W
被引数: 57
学者 查看更多机构
引用论文

引用论文

A chirp-compensated, injection-seeded alexandrite laser
err2014-01-23
err0
PREAI
errP. Bakule; P.E.G. Baird; M.G. Boshier; S.L. Cornish; D.F. Heller; K. Jungmann; I.C. Lane; V. Meyer; P.H.G. Sandars; W.T. Toner; M. Towrie; J.C. Walling
err分享
err收藏
Mutation Scanning Using MUT-MAP, a High-Throughput, Microfluidic Chip-Based, Multi-Analyte Panel
err2012-12-17
err0
errOAAI
errRajesh Patel; Alison Tsan; Rachel Tam; Rupal Desai; Nancy Schoenbrunner; Thomas W. Myers; Keith Bauer; Edward Smith; Rajiv Raja
err分享
err收藏
Use of the University of California Los Angeles Integrated Staging System to Predict Survival in Renal Cell Carcinoma: An International Multicenter Study
err2004-08-15
err0
PREAI
errJean-Jacques Patard; Hyung L. Kim; John S. Lam; Frederick J. Dorey; Allan J. Pantuck; Amnon Zisman; Vincenzo Ficarra; Ken-Ryu Han; Luca Cindolo; Alexandre De La Taille; Jacques Tostain; Walter Artibani; Colin P. Dinney; Christopher G. Wood; David A. Swanson; Claude C. Abbou; Bernard Lobel; Peter F.A. Mulders; Dominique K. Chopin; Robert A. Figlin; Arie S. Belldegrun
err分享
err收藏
Staging of Thyroid Cancer
err2006-01-01
err0
PREAI
errLeonard Wartofsky
err分享
err收藏
GPU computingGPU计算
err2008-05-01
err1.4K
PREAI
errOwens, John D.; Houston, Mike; Luebke, David; Green, Simon; Stone, John E.; Phillips, James C.
err分享
err收藏
err分享
err收藏
学者 查看更多内容