Self-Referencing Agents for Unsupervised Reinforcement Learning.
basic_science · Level V
Where this comes from
- Record sourced from PubMed, PMID 40198945.
- Also identified by DOI 10.1016/j.neunet.2025.107448.
- No licence information is recorded for this record.
- Because redistribution is not established, this page shows the abstract only. Follow the links below for the full text.
Abstract
Current unsupervised reinforcement learning methods often overlook reward nonstationarity during pre-training and the forgetting of exploratory behavior during fine-tuning. Our study introduces Self-Reference (SR), a novel add-on module designed to address both issues. SR stabilizes intrinsic rewards through historical referencing in pre-training, mitigating nonstationarity. During fine-tuning, it preserves exploratory behaviors, retaining valuable skills. Our approach significantly boosts the performance and sample efficiency of existing URL model-free methods on the Unsupervised Reinforcement Learning Benchmark, improving IQM by up to 17% and reducing the Optimality Gap by 31%. This highlights the general applicability and compatibility of our add-on module with existing methods.
Medical subject headings
- Reinforcement, Psychology
- Unsupervised Machine Learning
- Neural Networks, Computer