Related Experiment Video
Updated: Sep 12, 2025

The HoneyComb Paradigm for Research on Collective Human Behavior
Published on: January 19, 2019
Memory-Efficient Inverse Reinforcement Learning for Multiplayer Differential Games
None:
Data-driven inverse reinforcement learning (RL) control aims to infer the unknown cost function of a learner system from expert demonstrations. The convergence of existing methods necessitates a data storage mechanism to maintain persistent excitation (PE), which consumes memory and induces delays in satisfying full-rank conditions. To address these problems, in this article, we propose a novel memory-efficient inverse RL algorithm for multiplayer differential game that eliminates the need for strict PE and data storage. We prove that Nash equilibrium solutions for the learner system can be guaranteed under a mild initial excitation condition. Besides, existing inverse RL control algorithms often rely on an initial admissible control policy (IACP), which is difficult to obtain in data-driven scenarios. We address this problem by designing a novel filter-based homotopic RL algorithm, which derives an IACP for learner systems by shifting unstable poles into a stable region. Moreover, we establish several properties of the designed algorithms, including convergence, nonuniqueness, and stability. Finally, the effectiveness of the proposed algorithms is verified by comparative studies and simulation results.
Related Concept Videos
Reinforcement
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
Observational Learning
Reinforcement Schedules
Once a behavior is learned,...
Collisions in Multiple Dimensions: Problem Solving
A small car of mass 1,200 kg traveling east at 60 km/h collides at an intersection with a truck of mass 3,000 kg traveling due north at 40 km/h. The two vehicles are locked together. What is the...
Associative Learning
Classical conditioning, also known...
Dynamic Equilibrium

