Related Experiment Video
Updated: Jan 10, 2026

Investigating Motor Skill Learning Processes with a Robotic Manipulandum
Published on: February 12, 2017
SINDy-RL for interpretable and efficient model-based reinforcement learning
Nicholas Zolman1,2, Christian Lagemann3, Urban Fasel4
1Department of Mechanical Engineering, University of Washington, Seattle, WA, USA. nzolman@uw.edu.
This study introduces SINDy-RL, a new framework combining sparse dictionary learning and deep reinforcement learning (DRL). SINDy-RL creates efficient, interpretable control policies using significantly fewer training examples than traditional DRL.
Area of Science:
- Control Theory
- Machine Learning
- Fluid Dynamics
Background:
- Deep reinforcement learning (DRL) excels at complex control but demands extensive data and yields black-box policies.
- Sparse dictionary learning methods like SINDy offer efficient, interpretable models, particularly in low-data scenarios.
Purpose of the Study:
- To introduce SINDy-RL, a unified framework integrating SINDy and DRL.
- To develop efficient, interpretable, and trustworthy data-driven models for dynamics, rewards, and control policies.
- To address the data inefficiency and interpretability limitations of conventional DRL.
Main Methods:
- Integration of sparse identification of nonlinear dynamics (SINDy) with deep reinforcement learning (DRL).
- Development of a unifying framework (SINDy-RL) for learning dynamics, reward functions, and control policies.
- Application to benchmark control tasks and flow control problems, including gust mitigation on an airfoil.
Main Results:
- SINDy-RL achieves performance comparable to state-of-the-art DRL algorithms.
- The framework requires significantly fewer environmental interactions for training compared to traditional DRL.
- The resulting control policy is orders of magnitude smaller and more interpretable than DRL-derived policies.
Conclusions:
- SINDy-RL offers a more data-efficient and interpretable alternative to standard DRL for control tasks.
- The framework provides trustworthy and computationally efficient models suitable for various applications, including embedded systems.
- This approach enhances the practical applicability of reinforcement learning in complex dynamic environments.
Related Concept Videos
Reinforcement Schedules
Once a behavior is learned,...
Reinforcement
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
Observational Learning
Steps in the Modeling Process
Attention is the first necessary component for observational learning. It involves focusing on what the model is doing and saying. For example, if you decide to take a drawing class to enhance your skills, you need to pay close attention to the instructor's words and hand movements. The characteristics of the model significantly...
Comparison between RL and RC circuits
Stereotype Content Model

