Reinforcement
Reinforcement Schedules
Observational Learning
Decision Making: P-value Method
Associative Learning
Masking and Demasking Agents
You might also read
Articles linked to this work by shared authors, journal, and citation graph.
Updated: Jun 17, 2025

A Modified Lean and Release Technique to Emphasize Response Inhibition and Action Selection in Reactive Balance
Published on: March 19, 2020
This study introduces dynamic policy balance (DPB) and weighted entropy regularization (WER) to address imbalanced training in multiagent reinforcement learning (RL). These methods improve individual policy learning and exploration efficiency for better overall performance.
Area of Science:
Background:
Purpose of the Study:
Main Methods:
Main Results:
Conclusions: