Related Experiment Video
Updated: Sep 16, 2025

A Networked Desktop Virtual Reality Setup for Decision Science and Navigation Experiments with Multiple Participants
Published on: August 26, 2018
Risk-aware autonomous search and rescue with multiagent reinforcement learning
Aowabin Rahman1, Salman Shuvo1, Samrat Chatterjee2
1Optimization and Control Group, Pacific Northwest National Laboratory, Richland, Washington, USA.
Abstract:
Autonomous navigation in dynamic high-consequence environments, such as search and rescue (SAR) missions, often relies on multiagent robotic systems that need to learn and adapt to changing conditions. Adversarial risks can introduce further challenges in such a setting where an autonomous agent may exhibit deviations in their learned actions from training to testing. Moreover, the uncertain environment itself may also evolve with additional obstacles that can emerge during testing compared to conditions when algorithmic training of autonomous agents was performed. In this paper, we first focus on mathematically formulating the autonomous SAR problem via a risk-aware multiagent reinforcement learning approach. Thereafter, we design and implement numerical experiments to evaluate our approach under diverse hazard scenarios with a centralized training and decentralized testing paradigm. Finally, we discuss our results and steps for further research.
Related Concept Videos
Collisions in Multiple Dimensions: Problem Solving
A small car of mass 1,200 kg traveling east at 60 km/h collides at an intersection with a truck of mass 3,000 kg traveling due north at 40 km/h. The two vehicles are locked together. What is the...
Avoidance Learning and Learned Helplessness
Avoidance learning occurs when an organism learns that a specific behavior can prevent an unpleasant outcome. For example, a student who receives a bad grade may start studying harder to avoid future poor grades. This behavior persists even when the negative outcome is no longer present. Avoidance learning is powerful because it maintains behavior in the absence of the...
Reinforcement
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
Observational Learning
Masking and Demasking Agents
There are many masking agents, such as cyanide, fluoride, triethanolamine, thiourea, and 2,3-bis(sulfanyl)propan-1-ol (formerly 2,3-dimercapto-1-propanol), with the masking agent chosen based on...
Associative Learning
Classical conditioning, also known...

