Related Experiment Video
Updated: Aug 19, 2025

Real-Time Proxy-Control of Re-Parameterized Peripheral Signals using a Close-Loop Interface
Published on: May 8, 2021
Bayesian Disturbance Injection: Robust imitation learning of flexible policies for robot manipulation
Hanbit Oh1, Hikaru Sasaki1, Brendan Michael1
1Division of Information Science, Graduate School of Science and Technology, NAIST, 8916-5, Takayama-cho, Ikoma-city, 630-0192, Nara, Japan.
Abstract:
Humans demonstrate a variety of interesting behavioral characteristics when performing tasks, such as selecting between seemingly equivalent optimal actions, performing recovery actions when deviating from the optimal trajectory, or moderating actions in response to sensed risks. However, imitation learning, which attempts to teach robots to perform these same tasks from observations of human demonstrations, often fails to capture such behavior. Specifically, commonly used learning algorithms embody inherent contradictions between the learning assumptions (e.g., single optimal action) and actual human behavior (e.g., multiple optimal actions), thereby limiting robot generalizability, applicability, and demonstration feasibility. To address this, this paper proposes designing imitation learning algorithms with a focus on utilizing human behavioral characteristics, thereby embodying principles for capturing and exploiting actual demonstrator behavioral characteristics. This paper presents the first imitation learning framework, Bayesian Disturbance Injection (BDI), that typifies human behavioral characteristics by incorporating model flexibility, robustification, and risk sensitivity. Bayesian inference is used to learn flexible non-parametric multi-action policies, while simultaneously robustifying policies by injecting risk-sensitive disturbances to induce human recovery action and ensuring demonstration feasibility. Our method is evaluated through risk-sensitive simulations and real-robot experiments (e.g., table-sweep task, shaft-reach task and shaft-insertion task) using the UR5e 6-DOF robotic arm, to demonstrate the improved characterization of behavior. Results show significant improvement in task performance, through improved flexibility, robustness as well as demonstration feasibility.
Related Concept Videos
Observational Learning
Propagation of Uncertainty from Systematic Error
Rigid Body Equilibrium Problems - II
Consider two children sitting on a seesaw, which has negligible mass. The first child has a mass (m1) of 26 kg and sits at point A, which is 1.6 meters (r1) from the pivot point B; the second child has a mass (m2) of 32 kg and sits at point C. How far from the pivot point B should the second child sit (r2) to balance the seesaw?
Multi-input and Multi-variable systems
In the absence...
Propagation of Uncertainty from Random Error
Principle of Angular Impulse and Momentum: Problem Solving
Initially, a free-body diagram of the system is drawn to illustrate all the forces acting upon the system, providing a crucial understanding of the dynamics at play. Then, the principle of angular impulse and momentum is...

