Related Experiment Video
Updated: Jul 14, 2025

Recording Single Neurons' Action Potentials from Freely Moving Pigeons Across Three Stages of Learning
Published on: June 2, 2014
Exploring storm petrel pattering and sea-anchoring using deep reinforcement learning
Jiaqi Xue1,2,3, Fei Han1,2, Brett Klaassen van Oorschot4
1Key Laboratory of Coastal Environment and Resources of Zhejiang Province, School of Engineering, Westlake University, Hangzhou, Zhejiang 310030, People's Republic of China.
Abstract:
Developing hybrid aerial-aquatic vehicles that can interact with water surfaces while remaining aloft is valuable for various tasks, including ecological monitoring, water quality sampling, and search and rescue operations. Storm petrels are a group of pelagic seabirds that exhibit a unique locomotion pattern known as 'pattering' or 'sea-anchoring,' which is hypothesized to support forward locomotion and/or stationary posture at the water surface. In this study, we use morphological measurements of three storm petrel species and aero/hydrodynamic models to develop a computational storm petrel model and interact it with a hybrid fluid environment. Using deep reinforcement learning algorithms, we find that the storm petrel model exhibits high maneuverability and stability under a wide range of constant wind velocities after training. We also verify in the simulation that the storm petrel can use its 'pattering' or 'sea-anchoring' behavior to achieve different biomechanical sub-tasks (e.g. weight support, forward locomotion, stabilization) and adapt it under different wind speeds and optimization objectives. Specifically, we observe an adjustment in storm petrel's movement patterns as wind velocity increases and quantitively analyze its biomechanics underneath. Our results provide new insights into how storm petrels achieve efficient locomotion and dynamic stability at the air-water interface and adapt their behaviors to different wind velocities and tasks in open environments. Ultimately, our study will guide the design of next-generation biomimetic petrel-inspired robots for tasks requiring proximity to the water interface and efficiency.
Related Concept Videos
Reinforcement
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
Observational Learning
Reinforcement Schedules
Once a behavior is learned,...
Avoidance Learning and Learned Helplessness
Avoidance learning occurs when an organism learns that a specific behavior can prevent an unpleasant outcome. For example, a student who receives a bad grade may start studying harder to avoid future poor grades. This behavior persists even when the negative outcome is no longer present. Avoidance learning is powerful because it maintains behavior in the absence of the...
Role of Shaping in Operant Conditioning
The steps involved in shaping begin with reinforcing any response that resembles the desired behavior. For example, parents might praise a child for picking up one toy. As...
Associative Learning
Classical conditioning, also known...

