返回
Decoupled knowledge distillation method based on meta-learning
DOI:10.1016/j.hcc.2023.100164.png)
摘要
En 中文
With the advancement of deep learning techniques, the number of model parameters has been increasing, leading to significant memory consumption and limits in the deployment of such models in real-time applications. To reduce the number of model parameters and enhance the generalization capability of neural networks, we propose a method called Decoupled MetaDistil, which involves decoupled meta-distillation. This method utilizes meta-learning to guide the teacher model and dynamically adjusts the knowledge transfer strategy based on feedback from the student model, thereby improving the generalization ability. Furthermore, we introduce a decoupled loss method to explicitly transfer positive sample knowledge and explore the potential of negative samples knowledge. Extensive experiments demonstrate the effectiveness of our method. (c) 2023 The Author(s). Published by Elsevier B.V. on behalf of Shandong University. This is an open access article under the CC BY-NC-ND license (http://creativecommons.org/licenses/by-nc-nd/4.0/).
Keyword:
Model compression
Knowledge distillation
Meta-learning
Decoupled loss
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
H
IF:
3
论文数:
244
被引数:
407
机构
引用论文
Using the Unified Protocol for Transdiagnostic Treatment of Emotional Disorders With Youth Exhibiting Anger and Irritability使用统一的协议对表现出愤怒和易怒的年轻人进行情绪障碍的诊断治疗
没有更多内容

