arrow
返回

Parallel Blockwise Knowledge Distillation for Deep Neural Network Compression

delete2021-07-01
delete23
delete
OA
AI
C
Cody Blakeney
X
Xiaomin Li
Y
Yan Yan
Z
Ziliang Zong *
DOI:10.1109/TPDS.2020.3047003delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Deep neural networks (DNNs) have been extremely successful in solving many challenging AI tasks in natural language processing, speech recognition, and computer vision nowadays. However, DNNs are typically computation intensive, memory demanding, and power hungry, which significantly limits their usage on platforms with constrained resources. Therefore, a variety of compression techniques (e.g., quantization, pruning, and knowledge distillation) have been proposed to reduce the size and power consumption of DNNs. Blockwise knowledge distillation is one of the compression techniques that can effectively reduce the size of a highly complex DNN. However, it is not widely adopted due to its long training time. In this article, we propose a novel parallel blockwise distillation algorithm to accelerate the distillation process of sophisticated DNNs. Our algorithm leverages local information to conduct independent blockwise distillation, utilizes depthwise separable layers as the efficient replacement block architecture, and properly addresses limiting factors (e.g., dependency, synchronization, and load balancing) that affect parallelism. The experimental results running on an AMD server with four Geforce RTX 2080Ti GPUs show that our algorithm can achieve 3x speedup plus 19 percent energy savings on VGG distillation, and 3.5x speedup plus 29 percent energy savings on ResNet distillation, both with negligible accuracy loss. The speedup of ResNet distillation can be further improved to 3.87 when using four RTX6000 GPUs in a distributed cluster.
Keyword:
Computational modeling
Training
Task analysis
Quantization (signal)
Neural networks
Deep learning
Hardware
Deep neural networks
model compression
knowledge distillation
parallel training

期刊

IEEE Transactions on Parallel and Distributed Systems 封面图
IEEE Transactions on Parallel and Distributed Systems
IF:
6
论文数:
5.2K
被引数:
1.1W

机构

Texas State University System 封面图
Texas State University System
学者数:
5.5K
论文数: 4.8K
被引数: 13
引用论文

引用论文

err分享
err收藏
Resolution of symptoms in neuroleptic malignant syndrome
err2010-01-01
err0
errOAAI
errAshish Srivastava; YvonneD.S. Pereira; BramhanandS Cuncoliencar; Nayana Naik
err分享
err收藏
err分享
err收藏
Clozapine Use Presenting with Pseudopheochromocytoma in a Schizophrenic Patient: A Case Report
err2013-01-01
err0
errOAAI
errJaskanwal Sara; Matt Jenkins; Tanveer Chohan; Karan Jolly; Lisa Shepherd; Nirav Y. Gandhi; Jayadave Shakher
err分享
err收藏
ImageNet Large Scale Visual Recognition ChallengeImageNet大规模视觉识别挑战
err2015-04-11
err2.7W
PREAI
errRussakovsky, Olga; Deng, Jia; Su, Hao; Krause, Jonathan; Satheesh, Sanjeev; Ma, Sean; Huang, Zhiheng; Karpathy, Andrej; Khosla, Aditya; Bernstein, Michael; Berg, Alexander C.; Fei-Fei, Li
err分享
err收藏
Synthetic Depth-of-Field with a Single-Camera Mobile Phone
err2018-07-30
err125
errOAAI
errWadhwa, Neal; Garg, Rahul; Jacobs, David E.; Feldman, Bryan E.; Kanazawa, Nori; Carroll, Robert; Movshovitz-Attias, Yair; Barron, Jonathan T.; Pritch, Yael; Levoy, Marc
err分享
err收藏
Macrophage Migration Inhibitory Factor Plays a Critical Role in Mediating Protection against the Helminth ParasiteTaenia crassiceps
err2003-03-01
err0
errOAAI
errMiriam Rodríguez-Sosa; Lucia E. Rosas; John R. David; Rafael Bojalil; Abhay R. Satoskar; Luis I. Terrazas
err分享
err收藏
学者 查看更多内容