arrow
返回

A Stage-Level Network Parallelization Method Based on Depth Decomposition

delete2024-01-01
delete0
delete
OA
AI
Z
Zuming Wu
张
张云蔚 (Zhang, Yunwei) *
李
李斌 (Bin Li)
C
Chengjin Tao
DOI:10.1109/ACCESS.2024.3353221delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Parallel computational operations can significantly enhance network computational efficiency, and such processing has a wide range of applications across different spatial scales in networks. However, in the stage-wise level of networks, the majority still defaults to maintaining a serial single-chain structure. We propose a method of splitting across the depth dimension in multiple consecutive stacked computational units, extending such parallel computational operations to the stage-wise level of the network. This type of processing does not introduce excessive computational latency at the network stage level, and the additional computation primarily comes from the Fusion structures between different branches and the widened stage Transition layers. However, due to the relatively small proportion of this additional computational load compared to the entire network and its ease of maintenance, it is manageable. Compared to the performance of a single chain network scaled in depth, the introduction of parallel structures compresses its depth and diversifies the width of the network stage level. Improved the relationship between the initial network's accuracy and computational efficiency within a certain range. Through subsequent improvements to the stage transition layers and the introduction of branch attention, the performance of the parallelized structure can be further enhanced. In conclusion, our approach provides a viable practice for introducing parallel structures within the stage level of stage-wise networks. By transforming the original serial structure of continuously stacked computational units in the stage level into a parallel structure with multiple subnets, we can achieve superior overall performance compared to the original network within a certain range. The code is available at https://github.com/forrest996/ResNet_P.
Keyword:
Deep learning
neural networks
parallelization
computational efficiency

期刊

IEEE Access 封面图
IEEE Access
IF:
3.6
论文数:
9.8W
被引数:
29.4W

机构

C
China National Tobacco Corporation
学者数:
3.5K
论文数: 2.4K
被引数: 2
引用论文

引用论文

Synthesis of n-type semiconducting diamond film using diphosphorus pentaoxide as the doping source以五氧化二磷为掺杂源合成n型半导体金刚石膜
err1990-10-01
err0
PREAI
errKen Okano; Hideo Kiyota; Tatsuya Iwasaki; Yoshitaka Nakamura; Yukio Akiba; Tateki Kurosu; Masamori Iida; Terutaro Nakamura
err分享
err收藏
Toward Better Accuracy-Efficiency Trade-Offs: Divide and Co-Training
err2022-01-01
err14
errOAAI
errZhao, Shuai; Zhou, Liguang; Wang, Wenxiao; Cai, Deng; Lam, Tin Lun; Xu, Yangsheng
err分享
err收藏
学者 查看更多内容