arrow
返回

Distribution-Sensitive Information Retention for Accurate Binary Neural Network

delete2022-10-02
delete35
PRE
AI
H
Haotong Qin
X
Xiangguo Zhang
R
Ruihao Gong
Y
Yifu Ding
Y
Yi Xu
X
Xianglong Liu *
DOI:10.1007/s11263-022-01687-5delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Model binarization is an effective method of compressing neural networks and accelerating their inference process, which enables state-of-the-art models to run on resource-limited devices. Recently, advanced binarization methods have been greatly improved by minimizing the quantization error directly in the forward process. However, a significant performance gap still exists between the 1-bit model and the 32-bit one. The empirical study shows that binarization causes a great loss of information in the forward and backward propagation which harms the performance of binary neural networks (BNNs). We present a novel distribution-sensitive information retention network (DIR-Net) that retains the information in the forward and backward propagation by improving internal propagation and introducing external representations. The DIR-Net mainly relies on three technical contributions: (1) Information Maximized Binarization (1MB): minimizing the information loss and the binarization error of weights/activations simultaneously by weight balance and standardization; (2) Distribution-sensitive Tivo-stage Estimator (DTE): retaining the information of gradients by distribution-sensitive soft approximation by jointly considering the updating capability and accurate gradient; (3) Representation-align Binarization-aware Distillation (RBD): retaining the representation information by distilling the representations between full-precision and binarized networks. The DIR-Net investigates both forward and backward processes of BNNs from the unified information perspective, thereby providing new insight into the mechanism of network binarization. The three techniques in our DIR-Net are versatile and effective and can be applied in various structures to improve BNNs. Comprehensive experiments on the image classification and objective detection tasks show that our DIR-Net consistently outperforms the state-of-the-art binarization approaches under mainstream and compact architectures, such as ResNet, VGG, EfficientNet, DARTS, and MobileNet. Additionally, we conduct our DIR-Net on real-world resource-limited devices which achieves 11.1x storage saving and 5.4x speedup.
Keyword:
Binary neural network
Network quantization
Model compression
Deep learning

期刊

International Journal of Computer Vision 封面图
International Journal of Computer Vision
IF:
9.3
论文数:
3.9K
被引数:
2.8W

机构

B
Beihang University
学者数:
5.2W
论文数: 4.1W
被引数: 37
引用论文

引用论文

err分享
err收藏
Acoustoelectric Effect
err1959-01-01
err0
PREAI
errR. H. Parmenter
err分享
err收藏
Characterization of the Reaction Performance for Residue Hydrotreating Feedstocks
err2010-12-16
err0
PREAI
errYu-dong Sun; Chao-he Yang; Hong-hong Shan; Ben-xian Shen
err分享
err收藏
Unified Binary Generative Adversarial Network for Image Retrieval and Compression
err2020-02-18
err54
errOAAI
errSong, Jingkuan; He, Tao; Gao, Lianli; Xu, Xing; Hanjalic, Alan; Shen, Heng Tao
err分享
err收藏
Digital cavities and their potential applications
err2013-05-21
err0
errOAAI
errK Karki; M Torbjörnsson; J R Widom; A H Marcus; T Pullerits
err分享
err收藏
学者 查看更多内容