arrow
Return

Self-Referencing Agents for Unsupervised Reinforcement Learning

delete2025-05-27
delete0
PRE
AI
Z
Zhao, Andrew
E
Erle Zhu
R
Rui Lu
M
Matthieu Lin
Y
Yong‐Jin Liu
G
Gao Huang *
DOI:10.1016/j.neunet.2025.107448delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Current unsupervised reinforcement learning methods often overlook reward nonstationarity during pre-training and the forgetting of exploratory behavior during fine-tuning. Our study introduces Self-Reference (SR), a novel add-on module designed to address both issues. SR stabilizes intrinsic rewards through historical referencing in pre-training, mitigating nonstationarity. During fine-tuning, it preserves exploratory behaviors, retaining valuable skills. Our approach significantly boosts the performance and sample efficiency of existing URL model-free methods on the Unsupervised Reinforcement Learning Benchmark, improving IQM by up to 17% and reducing the Optimality Gap by 31%. This highlights the general applicability and compatibility of our add-on module with existing methods.
Keywords:
Reinforcement learning
Unsupervised reinforcement learning
Pretraining
Finetuning

Journal

Neural Networks cover
Neural Networks
IF:
6.3
Papers:
7.8K
Citations:
3.0W

Organization

No organization information available