arrow
Return

A Stage-Level Network Parallelization Method Based on Depth Decomposition

delete2024-01-01
delete0
delete
OA
AI
Z
Zuming Wu
张
张云蔚 (Zhang, Yunwei) *
李
李斌 (Bin Li)
C
Chengjin Tao
DOI:10.1109/ACCESS.2024.3353221delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Parallel computational operations can significantly enhance network computational efficiency, and such processing has a wide range of applications across different spatial scales in networks. However, in the stage-wise level of networks, the majority still defaults to maintaining a serial single-chain structure. We propose a method of splitting across the depth dimension in multiple consecutive stacked computational units, extending such parallel computational operations to the stage-wise level of the network. This type of processing does not introduce excessive computational latency at the network stage level, and the additional computation primarily comes from the Fusion structures between different branches and the widened stage Transition layers. However, due to the relatively small proportion of this additional computational load compared to the entire network and its ease of maintenance, it is manageable. Compared to the performance of a single chain network scaled in depth, the introduction of parallel structures compresses its depth and diversifies the width of the network stage level. Improved the relationship between the initial network's accuracy and computational efficiency within a certain range. Through subsequent improvements to the stage transition layers and the introduction of branch attention, the performance of the parallelized structure can be further enhanced. In conclusion, our approach provides a viable practice for introducing parallel structures within the stage level of stage-wise networks. By transforming the original serial structure of continuously stacked computational units in the stage level into a parallel structure with multiple subnets, we can achieve superior overall performance compared to the original network within a certain range. The code is available at https://github.com/forrest996/ResNet_P.
Keywords:
Deep learning
neural networks
parallelization
computational efficiency

Journal

IEEE Access cover
IEEE Access
IF:
3.6
Papers:
9.8W
Citations:
29.4W

Organization

C
China National Tobacco Corporation
Scholars:
3.5K
Papers: 2.4K
Citations: 2
Cited Papers

Cited Papers

errShare
errSave
Determination of the optical band-gap energy of cubic and hexagonal boron nitride using luminescence excitation spectroscopy
err2008-01-31
err0
errOAAI
errD A Evans; A G McGlynn; B M Towlson; M Gunn; D Jones; T E Jenkins; R Winter; N R J Poolton
errShare
errSave
Synthesis of n-type semiconducting diamond film using diphosphorus pentaoxide as the doping source
err1990-10-01
err0
PREAI
errKen Okano; Hideo Kiyota; Tatsuya Iwasaki; Yoshitaka Nakamura; Yukio Akiba; Tateki Kurosu; Masamori Iida; Terutaro Nakamura
errShare
errSave
Energy-Efficient Robot Configuration and Motion Planning Using Genetic Algorithm and Particle Swarm Optimization
err2022-03-11
err0
errOAAI
errKazuki Nonoyama; Ziang Liu; Tomofumi Fujiwara; Md Moktadir Alam; Tatsushi Nishi
errShare
errSave
Structural study of lanthanides(III) in aqueous nitrate and chloride solutions by EXAFS
err1999-02-01
err0
PREAI
errT. Yaita; H. Narita; Sh. Suzuki; Sh. Tachimori; H. Motohashi; H. Shiwaku
errShare
errSave
Toward Better Accuracy-Efficiency Trade-Offs: Divide and Co-Training
err2022-01-01
err14
errOAAI
errZhao, Shuai; Zhou, Liguang; Wang, Wenxiao; Cai, Deng; Lam, Tin Lun; Xu, Yangsheng
errShare
errSave
researcher View more