arrow
Return

A framework for scheduling dependent programs on GPU architectures

delete2020-06-01
delete5
PRE
AI
Y
Yuan‐Ming Chang
W
Wei-Cheng Liao
S
Shao-Chung Wang
C
Chun‐Chieh Yang
DOI:10.1016/j.sysarc.2020.101712delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
In recent years, the parallel computing performance on GPUs (graphics processing units) has grown rapidly. As a result, GPUs have been widely applied in computationally intensive applications such as image processing, deep learning, and artificial intelligence. Because these applications can be modeled by multiple GPU kernels, some of which might even be dependent, it is essential to identify an efficient method for scheduling dependent kernels on GPU cores. Simply observing kernel dependencies by executing them in sequence results in performance degradation. Furthermore, dependent kernels generally need to share data. Consequently, without properly scheduling dependent kernels, unnecessary memory accesses and copies will be generated. Neural network model environments include many operators that are suitable for both parallel and dependent kernels. This paper proposes an efficient and clear method for creating a framework to analyze the ONNX (open neural network exchange) model and find the pattern of dependent kernels, which can then be used for scheduling with the GPU architecture. The preliminary experimental results show that this technique improves the overall performance by 8% and reduces the cache miss rate by 14% on average by combining neural network operators and appropriate memory policies.
Keywords:
GPU
ONNX
Scheduling
Memory
Simulator
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Journal of Systems Architecture cover
Journal of Systems Architecture
IF:
4.1
Papers:
3.0K
Citations:
4.2K

Organization

N
National Tsing Hua University
Scholars:
1.6W
Papers: 1.4W
Citations: 1.7W
N
national taiwan university of science & technology
Scholars:
8.8K
Papers: 8.7K
Citations: 9
Cited Papers

Cited Papers

Enabling PoCL-based runtime frameworks on the HSA for OpenCL 2.0 support
err2017-11-01
err8
PREAI
errChang, Yuan-Ming; Wang, Shao-Chung; Yang, Chun-Chieh; Hwang, Yuan-Shin; Lee, Jenq-Kuen
errShare
errSave
Direct Measurement of the Spin-Orbit Interaction in a Two-Electron InAs Nanowire Quantum Dot
err2007-06-26
err0
errOAAI
errC. Fasth; A. Fuhrer; L. Samuelson; Vitaly N. Golovach; Daniel Loss
errShare
errSave