arrow
返回

Progressive Bitwidth Assignment Approaches for Efficient Capsule Networks Quantization

delete2025-01-01
delete0
delete
OA
AI
M
Mohsen Raji *
A
Amir Ghazizadeh Ahsaei
K
Kimia Soroush
B
Behnam Ghavami
DOI:10.1109/ACCESS.2025.3534434delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Capsule Networks (CapsNets) are a class of neural network architectures that can be used to more accurately model hierarchical relationships due to their hierarchical structure and dynamic routing algorithms. However, their high accuracy comes at the cost of significant memory and computational resources, making them less feasible for deployment on resource-constrained devices. In this paper, progressive bitwidth assignment approaches are introduced to efficiently quantize the CapsNets. Initially, a comprehensive and detailed analysis of parameter quantization in CapsNets is performed exploring various granularities, such as block-wise quantization and dynamic routing quantization. Then, three quantization approaches are applied to progressively quantize the CapsNet, considering various insights into the susceptibility of layers to quantization. The proposed approaches include Post-Training Quantization (PTQ) strategies that minimize the dependence on floating-point operations and incorporates layer-specific integer bit-widths based on quantization error analysis. PTQ strategies employ Power-of-Two (PoT) scaling factors to simplify computations, effectively utilizing hardware shifts and significantly reducing the computational complexity. This technique not only reduces the memory footprint but also maintains accuracy by introducing a range clipping method tailored to the hardware's capabilities, obviating the need for data preprocessing. Our experimental results on ShallowCaps and DeepCaps networks across multiple datasets (MNIST, Fashion-MNIST, CIFAR-10, and SVHN) demonstrate the efficiency of our approach. Specifically, on the CIFAR-10 dataset using the DeepCaps architecture, we achieved a substantial memory reduction (7.02x for weights and 3.74x for activations) with a minimal accuracy loss of only 0.09%. By using progressive bitwidth assignment and post-training quantization, this work optimizes CapsNets for efficient, real-time visual processing on resource-constrained edge devices, enabling applications in IoT, mobile platforms, and embedded systems.
Keyword:
Capsule networks
deep learning
neural networks
post-training quantization
compression
Capsule networks
deep learning
neural networks
post-training quantization
compression

期刊

IEEE Access 封面图
IEEE Access
IF:
3.6
论文数:
9.8W
被引数:
29.4W

机构

S
shahid bahonar university of kerman (sbuk)
学者数:
3.2K
论文数: 3.0K
被引数: 0
S
Shiraz University
学者数:
8.1K
论文数: 7.5K
被引数: 7.4K
引用论文

引用论文

err
IF0
err
err0
PREAI
err
err分享
err收藏
Deep Learning for AIAI的深度学习
err2021-06-21
err312
errOAAI
errBengio, Yoshua; Lecun, Yann; Hinton, Geoffrey
err分享
err收藏
Residual Quantization for Low Bit-Width Neural Networks低位宽神经网络的残差量化
err2023-01-01
err6
PREAI
errLi, Zefan; Ni, Bingbing; Yang, Xiaokang; Zhang, Wenjun; Gao, Wen
err分享
err收藏
err分享
err收藏
Three-dimensional reconstruction
err1992-12-01
err0
PREAI
errM. Vahlensieck; Ph. Lang; W. P. Chang; S. Grampp; H. K. Genant
err分享
err收藏
Where do winds come from? A new theory on how water vapor condensation influences atmospheric pressure and dynamics
err
IF0
err2010-10-15
err0
errOAAI
errA. M. Makarieva; V. G. Gorshkov; D. Sheil; A. D. Nobre; B.-L. Li
err分享
err收藏
学者 查看更多内容