arrow
返回

Knowledge fusion distillation and gradient-based data distillation for class-incremental learning

delete2025-03-01
delete0
PRE
AI
L
Lin Xiong
X
Xin Guan
H
Hailing Xiong *
K
Kangwen Zhu
F
Fuqing Zhang
DOI:10.1016/j.neucom.2024.129286delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Deep neural networks, despite their exceptional performance on individual tasks, face the challenge of catastrophic forgetting when incrementally learning new scenarios. Existing class-incremental learning methods have shown promising results by rehearsing past samples from auxiliary memory. However, preserving only a limited subset of previous samples fails to capture the complete distribution, leading to a decline in memory quality as new data is added. Additionally, issues like exemplar distribution collapse and task-recency bias impede the effective transfer of knowledge from teacher to student during distillation. To address these problems, we propose a novel replay-based learning framework called Knowledge fusion Distillation and gradient-based Data Distillation (K3D). K3D simultaneously optimizes both the classifier and exemplars during the learning phase. By parameterizing exemplars, we enable their optimization through gradient-based data distillation, enhancing their representativeness instead of discarding them, as done in other replay strategies. Furthermore, we enhance the classifier using knowledge fusion distillation, ensuring that decision boundaries are maintained across all classes. Extensive experiments on the CIFAR100 and ImageNet100 benchmarks show that synthetic exemplars optimized by K3D are more representative than selected ones. By combining optimizable exemplars and knowledge fusion distillation, K3D outperforms several state-of-the-art methods and effectively mitigates catastrophic forgetting in class-incremental learning scenarios.
Keyword:
Gradient-based data distillation
Optimizable exemplars
Synthetic replay
Knowledge fusion distillation
Class-incremental learning

期刊

Neurocomputing 封面图
Neurocomputing
IF:
6.5
论文数:
2.5W
被引数:
6.5W

机构

C
Chongqing Jiaotong University
学者数:
6.5K
论文数: 4.3K
被引数: 94
S
southwest university - china
学者数:
2.6W
论文数: 1.9W
被引数: 21
C
chinese academy of sciences
学者数:
56.7W
论文数: 45.0W
被引数: 704
学者 查看更多机构
引用论文

引用论文

ImageNet Large Scale Visual Recognition ChallengeImageNet大规模视觉识别挑战
err2015-04-11
err2.7W
PREAI
errRussakovsky, Olga; Deng, Jia; Su, Hao; Krause, Jonathan; Satheesh, Sanjeev; Ma, Sean; Huang, Zhiheng; Karpathy, Andrej; Khosla, Aditya; Bernstein, Michael; Berg, Alexander C.; Fei-Fei, Li
err分享
err收藏
An examination of breastmilk composition among high altitude Peruvian women
err2020-03-13
err0
PREAI
errLauren A. Schafrank; Jennifer R. Washabaugh; Morgan K. Hoke
err分享
err收藏
PyCIL: a Python toolbox for class-incremental learning
err2023-04-19
err19
errOAAI
errZhou, Da-Wei; Wang, Fu-Yun; Ye, Han-Jia; Zhan, De-Chuan
err分享
err收藏
学者 查看更多内容