Related Experiment Video
Updated: May 15, 2025

Investigating Motor Skill Learning Processes with a Robotic Manipulandum
Published on: February 12, 2017
Self-Referencing Agents for Unsupervised Reinforcement Learning
Andrew Zhao1, Erle Zhu2, Rui Lu1
1Department of Automation, BNRist, Tsinghua University, China.
Abstract:
Current unsupervised reinforcement learning methods often overlook reward nonstationarity during pre-training and the forgetting of exploratory behavior during fine-tuning. Our study introduces Self-Reference (SR), a novel add-on module designed to address both issues. SR stabilizes intrinsic rewards through historical referencing in pre-training, mitigating nonstationarity. During fine-tuning, it preserves exploratory behaviors, retaining valuable skills. Our approach significantly boosts the performance and sample efficiency of existing URL model-free methods on the Unsupervised Reinforcement Learning Benchmark, improving IQM by up to 17% and reducing the Optimality Gap by 31%. This highlights the general applicability and compatibility of our add-on module with existing methods.
Related Concept Videos
Observational Learning
Reinforcement Schedules
Once a behavior is learned,...
Self-Evaluation: Self-Enhancement and Self-Verification
Self-Presentation: Self-Monitoring and Self-Handicapping
Self-Schemas
Nonconscious Mimicry

