返回
Drifting explanations in continual learning
DOI:10.1016/j.neucom.2024.127960.png)
摘要
En 中文
Continual Learning (CL) trains models on streams of data, with the aim of learning new information without forgetting previous knowledge. However, many of these models lack interpretability, making it difficult to understand or explain how they make decisions. This lack of interpretability becomes even more challenging given the non-stationary nature of the data streams in CL. Furthermore, CL strategies aimed at mitigating forgetting directly impact the learned representations. We study the behavior of different explanation methods in CL and propose CLEX (ContinuaL EXplanations), an evaluation protocol to robustly assess the change of explanations in Class-Incremental scenarios, where forgetting is pronounced. We observed that models with similar predictive accuracy do not generate similar explanations. Replay-based strategies, well-known to be some of the most effective ones in class-incremental scenarios, are able to generate explanations that are aligned to the ones of a model trained offline. On the contrary, naive fine-tuning often results in degenerate explanations that drift from the ones of an offline model. Finally, we discovered that even replay strategies do not always operate at best when applied to fully-trained recurrent models. Instead, randomized recurrent models (leveraging on an untrained recurrent component) clearly reduce the drift of the explanations. This discrepancy between fully-trained and randomized recurrent models, previously known only in the context of their predictive continual performance, is more general, including also continual explanations.
Keyword:
Continual learning
Explainable AI
Lifelong learning
Recurrent neural networks
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
6.5
论文数:
2.5W
被引数:
6.5W
机构
引用论文
Continual learning for robotics: Definition, framework, learning strategies, opportunities and challenges机器人持续学习: 定义、框架、学习策略、机遇与挑战
INFORMATION FUSION
IF15.5
Continual learning for recurrent neural networks: An empirical evaluation递归神经网络的持续学习: 经验评估
NEURAL NETWORKS
IF6.3
Mutagenesis of Fujinami Sarcoma Virus: Evidence that tyrosine phosphorylation of P130gag-fps modulates its biological activity
Cell
IF0

