Related Experiment Video
Updated: Jan 28, 2026

C. elegans Positive Butanone Learning, Short-term, and Long-term Associative Memory Assays
Published on: March 11, 2011
LSWM: A Long-Short History World Model for Bipedal Locomotion via Reinforcement Learning
Jie Xue1,2,3, Zhiyuan Liang1,2,3, Haiming Mou3
1School of Optical-Electrical and Computer Engineering, University of Shanghai for Science and Technology, Shanghai 200093, China.
Abstract:
The presence of sensor noise, missing states and inadequate future prediction capabilities imposes significant limitations on the locomotion performance of bipedal robots operating in unstructured terrain. Conventional methods generally depend on long-term history observations to reconstruct single-frame privileged information. However, these methods fail to acknowledge the pivotal function of short-term history in rapid state responses and the significance of future state prediction in anticipating potential risks. The proposed framework is a Long-Short World Model (LSWM), which integrates state reconstruction and future state prediction to enhance the locomotion capabilities of bipedal robots in complex environments. The LSWM framework comprises two modules: a state reconstruction module (SRM) and a future state prediction module (SPM). The state reconstruction module employs long-term history observations to reconstruct privileged information in the current short-term history, thereby effectively improving the system's robustness to sensor noise and enhancing state observability. The future state prediction module enhances the robot's adaptability to complex environments and unpredictable scenarios by predicting the robot's future short-term privileged information. We conducted extensive comparative experiments in simulation as well as in a variety of real-world indoor and outdoor environments. In the indoor stair-climbing task, LSWM achieved a 94% success rate, outperforming the current state-of-the-art baseline methods by at least 34%, thereby demonstrating its substantial performance advantages in complex and dynamic environments.
More Related Videos
09:23JenaTron - An Experimental Approach to Study the Effects of Plant History and Soil History on Grassland Ecosystem Functioning
Published on: March 21, 2025
10:43Eye-tracking Technology and Data-mining Techniques used for a Behavioral Analysis of Adults engaged in Learning Processes
Published on: June 10, 2021
Related Concept Videos
What is Evolutionary History?
Reinforcement
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
History of Microbiology
Life Histories
Corrosion of Reinforcement
However, over time and under certain conditions like carbonation, chloride ingress, and cracking this protective state can be compromised. Steel has areas with...
Reinforcement Schedules
Once a behavior is learned,...