arrow
返回

SLAPP: Subgraph-level attention-based performance prediction for deep learning models

delete2024-02-01
delete2
PRE
AI
Z
Zhenyi Wang
P
Pengfei Yang *
L
Linwei Hu
B
Bowen Zhang
C
Cheng-Min Lin
W
Wenkai Lv
Q
Quan Wang
DOI:10.1016/j.neunet.2023.11.043delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
The intricacy of the Deep Learning (DL) landscape, brimming with a variety of models, applications, and platforms, poses considerable challenges for the optimal design, optimization, or selection of suitable DL models. One promising avenue to address this challenge is the development of accurate performance prediction methods. However, existing methods reveal critical limitations. Operator-level methods, proficient at predicting the performance of individual operators, often neglect broader graph features, which results in inaccuracies in full network performance predictions. On the contrary, graph-level methods excel in overall network prediction by leveraging these graph features but lack the ability to predict the performance of individual operators. To bridge these gaps, we propose SLAPP, a novel subgraph-level performance prediction method. Central to SLAPP is an innovative variant of Graph Neural Networks (GNNs) that we developed, named the Edge Aware Graph Attention Network (EAGAT). This specially designed GNN enables superior encoding of both node and edge features. Through this approach, SLAPP effectively captures both graph and operator features, thereby providing precise performance predictions for individual operators and entire networks. Moreover, we introduce a mixed loss design with dynamic weight adjustment to reconcile the predictive accuracy between individual operators and entire networks. In our experimental evaluation, SLAPP consistently outperforms traditional approaches in prediction accuracy, including the ability to handle unseen models effectively. Moreover, when compared to existing research, our method demonstrates a superior predictive performance across multiple DL models.
Keyword:
Deep Learning (DL)
Graph neural networks (GNNs)
Performance prediction
Computation graph optimization
Attention mechanisms

期刊

Neural Networks 封面图
Neural Networks
IF:
6.3
论文数:
8.2K
被引数:
3.0W

机构

X
Xidian University
学者数:
2.4W
论文数: 1.9W
被引数: 9.7K
引用论文

引用论文

ANNETTE: Accurate Neural Network Execution Time Estimation With Stacked Models
err2021-01-01
err12
errOAAI
errWess, Matthias; Ivanov, Matvey; Unger, Christoph; Nookala, Anvesh; Wendt, Alexander; Jantsch, Axel
err分享
err收藏
err分享
err收藏
err分享
err收藏
Efficient network architecture search via multiobjective particle swarm optimization based on decomposition
err2020-03-01
err59
PREAI
errJiang, Jing; Han, Fei; Ling, Qinghua; Wang, Jie; Li, Tiange; Han, Henry
err分享
err收藏
Evolution of Organic Aerosols in the Atmosphere
err2009-12-11
err0
PREAI
errJ. L. Jimenez; M. R. Canagaratna; N. M. Donahue; A. S. H. Prevot; Q. Zhang; J. H. Kroll; P. F. DeCarlo; J. D. Allan; H. Coe; N. L. Ng; A. C. Aiken; K. S. Docherty; I. M. Ulbrich; A. P. Grieshop; A. L. Robinson; J. Duplissy; J. D. Smith; K. R. Wilson; V. A. Lanz; C. Hueglin; Y. L. Sun; J. Tian; A. Laaksonen; T. Raatikainen; J. Rautiainen; P. Vaattovaara; M. Ehn; M. Kulmala; J. M. Tomlinson; D. R. Collins; M. J. Cubison; J. Dunlea; J. A. Huffman; T. B. Onasch; M. R. Alfarra; P. I. Williams; K. Bower; Y. Kondo; J. Schneider; F. Drewnick; S. Borrmann; S. Weimer; K. Demerjian; D. Salcedo; L. Cottrell; R. Griffin; A. Takami; T. Miyoshi; S. Hatakeyama; A. Shimono; J. Y Sun; Y. M. Zhang; K. Dzepina; J. R. Kimmel; D. Sueper; J. T. Jayne; S. C. Herndon; A. M. Trimborn; L. R. Williams; E. C. Wood; A. M. Middlebrook; C. E. Kolb; U. Baltensperger; D. R. Worsnop
err分享
err收藏
学者 查看更多内容