arrow
Return

Memory efficient data-free distillation for continual learning

delete2023-12-01
delete4
PRE
AI
X
Xiaorong Li
S
Shipeng Wang
孙
孙剑 (Jian Sun) *
Z
Zongben Xu
DOI:10.1016/j.patcog.2023.109875delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Deep neural networks suffer from the catastrophic forgetting phenomenon when trained on sequential tasks in continual learning, especially when data from previous tasks are unavailable. To mitigate catastrophic forgetting, various methods either store data from previous tasks, which may raise privacy concerns, or require large memory storage. Particularly, the distillation-based methods mitigate catastrophic forgetting by using proxy datasets. However, proxy datasets may not match the distributions of the original datasets of previous tasks. To address these problems in a setting where the full training data of previous tasks are unavailable and memory resources are limited, we propose a novel data-free distillation method. Our method encodes knowledge of previous tasks into network parameter gradients by Taylor expansion, deducing a regularizer relying on gradients in network training loss. To improve memory efficiency, we design an approach to compressing the gradients in the regularizer. Moreover, we theoretically analyze the approximation error of our method. Experimental results on multiple datasets demonstrate that our proposed method outperforms the existing approaches in continual learning.
Keywords:
Continual learning
Catastrophic forgetting
Knowledge distillation

Journal

Pattern Recognition cover
Pattern Recognition
IF:
7.6
Papers:
1.3W
Citations:
4.5W

Organization

X
xi'an jiaotong university
Scholars:
9.3W
Papers: 6.7W
Citations: 75
Cited Papers

Cited Papers

Exemplar-free class incremental learning via discriminative and comparable parallel one-class classifiers
err2023-08-01
err14
PREAI
errSun, Wenju; Li, Qingyong; Zhang, Jing; Wang, Danyu; Wang, Wen; Geng, YangLi-ao
errShare
errSave
FoCL: Feature-oriented continual learning for generative models
err2021-12-01
err9
errOAAI
errLao, Qicheng; Mortazavi, Mehrzad; Tahaei, Marzieh; Dutil, Francis; Fevens, Thomas; Havaei, Mohammad
errShare
errSave
Incremental Zero-Shot Learning
err2022-12-01
err35
PREAI
errWei, Kun; Deng, Cheng; Yang, Xu; Tao, Dacheng
errShare
errSave
SATS: Self-attention transfer for continual semantic segmentation
err2023-06-01
err14
errOAAI
errQiu, Yiqiao; Shen, Yixing; Sun, Zhuohao; Zheng, Yanchong; Chang, Xiaobin; Zheng, Weishi; Wang, Ruixuan
errShare
errSave
Continual learning of context-dependent processing in neural networks
err2019-08-09
err156
PREAI
errZeng, Guanxiong; Chen, Yang; Cui, Bo; Yu, Shan
errShare
errSave
Multi-criteria Selection of Rehearsal Samples for Continual Learning
err2022-12-01
err20
PREAI
errZhuang, Chen; Huang, Shaoli; Cheng, Gong; Ning, Jifeng
errShare
errSave
Overcoming catastrophic forgetting in neural networks
err2017-03-14
err3.7K
errOAAI
errKirkpatricka, James; Pascanu, Razvan; Rabinowitz, Neil; Veness, Joel; Desjardins, Guillaume; Rusu, Andrei A.; Milan, Kieran; Quan, John; Ramalho, Tiago; Grabska-Barwinska, Agnieszka; Hassabis, Demis; Clopath, Claudia; Kumaran, Dharshan; Hadsell, Raia
errShare
errSave
no more