arrow
返回

Relation Knowledge Distillation by Auxiliary Learning for Object Detection

delete2024-01-01
delete0
PRE
AI
H
Hao Wang
贾
贾同 (Tong Jia) *
Q
Qilong Wang
左
左旺孟 (Wangmeng Zuo)
DOI:10.1109/TIP.2024.3445740delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Balancing the trade-off between accuracy and speed for obtaining higher performance without sacrificing the inference time is a challenging topic for object detection task. Knowledge distillation, which serves as a kind of model compression techniques, provides a potential and feasible way to handle above efficiency and effectiveness issue through transferring the dark knowledge from the sophisticated teacher detector to the simple student one. Despite demonstrating promising solutions to make harmonies between accuracy and speed, current knowledge distillation for object detection methods still suffer from two limitations. Firstly, most of the methods are inherited or refereed from the frameworks in image classification task, and deploy an implicit manner by imitating or constraining the features from the intermediate layers or the output predictions between the teacher and student models. While little consideration has been raised to the intrinsic relevance of the classification and localization predictions in object detection task. Besides, these methods fail to investigate the relationship between detection and distillation tasks in knowledge distillation pipeline, and they train the whole network by simply integrating losses from these two different tasks through hand-crafted designation parameters. For addressing the aforementioned issues, we propose a novel Relation Knowledge Distillation by Auxiliary Learning for Object Detection (ReAL) method in this paper. Specifically, we first design a prediction relation distillation module which makes the student model directly mimic the output predictions from the teacher one, and conduct self and mutual relation distillation losses to excavate the relation information between teacher and student models. Moreover, for better devolving into the relationship between different tasks in distillation pipeline, we introduce the auxiliary learning into knowledge distillation for object detection and develop a dynamic weight adaptation strategy. Through regarding detection task as primary task and treating distillation task as auxiliary task in auxiliary learning framework, we dynamically adjust and regularize the corresponding weights of the losses for these tasks during the training process. Experiments on MS COCO dataset are conducted using various detector combinations of teacher and student models and the results show that our proposed ReAL can achieve obvious improvement on different distillation model configurations, while performing favorably against state-of-the-arts.
Keyword:
Task analysis
Predictive models
Object detection
Location awareness
Adaptation models
Accuracy
Head
knowledge distillation
relation information
auxiliary learning

期刊

IEEE Transactions on Image Processing 封面图
IEEE Transactions on Image Processing
IF:
13.7
论文数:
1.0W
被引数:
8.4W

机构

H
harbin institute of technology
学者数:
8.0W
论文数: 6.6W
被引数: 66
T
tianjin university
学者数:
8.0W
论文数: 5.8W
被引数: 88
N
northeastern university - china
学者数:
3.2W
论文数: 2.7W
被引数: 37
学者 查看更多机构
引用论文

引用论文

Conductivity Enhancement in Thin Silicon-on-Insulator Layer Embedding Artificial Dislocation Network
err2011-02-01
err0
PREAI
errYasuhiko Ishikawa; Kazuaki Yamauchi; Chihiro Yamamoto; Michiharu Tabe
err分享
err收藏
err分享
err收藏
Maternal obesity programs reduced leptin signaling in the pituitary and altered GH/IGF1 axis function leading to increased adiposity in adult sheep offspring
err2017-08-03
err0
errOAAI
errNuermaimaiti Tuersunjiang; John F. Odhiambo; Desiree R. Shasa; Ashley M. Smith; Peter W. Nathanielsz; Stephen P. Ford
err分享
err收藏
Synthesis of n-type semiconducting diamond film using diphosphorus pentaoxide as the doping source以五氧化二磷为掺杂源合成n型半导体金刚石膜
err1990-10-01
err0
PREAI
errKen Okano; Hideo Kiyota; Tatsuya Iwasaki; Yoshitaka Nakamura; Yukio Akiba; Tateki Kurosu; Masamori Iida; Terutaro Nakamura
err分享
err收藏
学者 查看更多内容