arrow
返回

Analyzing GCN Aggregation on GPU

delete2022-01-01
delete1
delete
OA
AI
I
Inje Kim
J
Jong Hyun Jeong
Y
Yunho Oh
M
Myung Kuk Yoon
G
Gunjae Koo *
DOI:10.1109/ACCESS.2022.3217222delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Graph convolutional neural networks (GCNs) are emerging neural networks for graph structures that include large features associated with each vertex. The operations of GCN can be divided into two phases - aggregation and combination. While the combination just performs matrix multiplications using trained weights and aggregated features, the aggregation phase requires graph traversal to collect features from adjacent vertices. Even though neural network applications rely on GPU's massively parallel processing, GCN aggregation kernels exhibit rather low performance since graph processing using compressed graph structures provokes frequent irregular accesses in GPUs. In order to investigate the performance hurdles of GCN aggregation on GPU, we perform an in-depth analysis of the aggregation kernels using real GPU hardware and a cycle-accurate GPU simulator. We first analyze the characteristics of the popular graph datasets used for GCN studies. We reveal the fractions of non-zero elements in feature vectors are diverse among datasets. Based on the observation, we build two types of aggregation kernels that handle uncompressed and compressed feature vectors. Our evaluation exhibits the performance of aggregation can be significantly influenced by kernel design approaches and feature density. We also analyze the individual loads that access the data arrays of the aggregation kernels to specify critical loads. Our analysis reveals the performance of GPU memory hierarchy is influenced by access patterns and feature size of graph datasets. Based on our observations we discuss possible kernel design approaches and architectural ideas that can improve the performance of GCN aggregation.
Keyword:
Graphics processing units
Kernel
Convolutional neural networks
Neural networks
Mathematical models
Hardware
Data models
Graph neural networks
GCN
aggregation kernel
GPU
characteristics

期刊

IEEE Access 封面图
IEEE Access
IF:
3.6
论文数:
9.8W
被引数:
29.4W

机构

K
Korea University
学者数:
3.6W
论文数: 3.8W
被引数: 4.4W
E
Ewha Womans University
学者数:
1.2W
论文数: 1.1W
被引数: 1.2W
引用论文

引用论文

A MORN1-associated HAD phosphatase in the basal complex is essential forToxoplasma gondiidaughter budding
err2016-03-09
err0
errOAAI
errKlemens Engelberg; F. Douglas Ivey; Angela Lin; Maya Kono; Alexander Lorestani; Dave Faugno-Fusci; Tim-Wolf Gilberger; Michael White; Marc-Jan Gubbels
err分享
err收藏
err
IF0
err
err0
PREAI
err
err分享
err收藏
err分享
err收藏
Cantú syndrome: Findings from 74 patients in the International Cantú Syndrome Registry
err2019-12-11
err0
errOAAI
errDorothy K. Grange; Helen I. Roessler; Conor McClenaghan; Karen Duran; Kathleen Shields; Maria S. Remedi; Nine V. A. M. Knoers; Jin‐Moo Lee; Edwin P. Kirk; Ingrid Scurr; Sarah F. Smithson; Gautam K. Singh; Mieke M. van Haelst; Colin G. Nichols; Gijs van Haaften
err分享
err收藏
Hemifield asymmetry in the potency of exogenous auditory and visual cues
err2011-06-01
err0
errOAAI
errYamaya Sosa; Aaron M. Clarke; Mark E. McCourt
err分享
err收藏
Checklist of the ichthyofauna of the Rio Negro basin in the Brazilian Amazon
err2019-10-17
err0
errOAAI
errHélio Beltrão; Jansen Zuanon; Efrem Ferreira
err分享
err收藏
学者 查看更多内容