You might also read
Articles linked to this work by shared authors, journal, and citation graph.
Updated: Aug 20, 2025

Quantifying Learning in Young Infants: Tracking Leg Actions During a Discovery-learning Task
Published on: June 1, 2015
This study introduces a new hybrid and dynamic policy gradient (HDPG) method for bipedal robot locomotion. HDPG improves control by optimizing multiple reward criteria simultaneously, outperforming traditional summed-reward deep reinforcement learning approaches.
Area of Science:
Background:
Purpose of the Study:
Main Methods:
Main Results:
Conclusions: