Q-learning with temporal memory to navigate turbulence

Marco Rando1, Martin James2, Alessandro Verri1

  • 1MaLGa, Department of computer science, bioengineering, robotics and systems engineering, University of Genova, Genova, Italy.

Arxiv
|May 7, 2024
PubMed
Summary

Agents can learn to navigate using only smell in turbulent environments. A reinforcement learning model with memory successfully guides agents by identifying key odor features and employing a crosswind search strategy, similar to insects.