Related Experiment Video
Updated: Dec 14, 2025

08:59
An Open-Source Virtual Reality System for the Measurement of Spatial Learning in Head-Restrained Mice
Published on: March 3, 2023
2.5K
Lessons from reinforcement learning for biological representations of space
Alex Muryy1, N Siddharth2, Nantas Nardelli2
1School of Psychology and Clinical Language Sciences, University of Reading, UK.
Vision Research
|July 20, 2020
Summary
This study explores reinforcement learning for spatial navigation, suggesting non-Cartesian representations may offer alternatives to traditional cognitive maps. These methods show promise for understanding biological spatial perception and navigation.
Area of Science:
- Neuroscience
- Artificial Intelligence
- Cognitive Science
Background:
- Neuroscience traditionally models spatial representation using various coordinate frames.
- Reinforcement learning (RL) offers a novel approach to spatial perception and navigation.
- Current RL models often avoid building explicit 3D 'maps'.
Purpose of the Study:
- To investigate RL methods that use image-based rewards for spatial tasks.
- To evaluate the geometric consistency of learned spatial representations.
- To explore alternatives to the 'cognitive map' theory in neuroscience.
Main Methods:
- Focus on RL agents rewarded for reaching target images.
- Testing geometric consistency through interpolation of learned locations.
- Introducing a hand-crafted, geometrically consistent representation.
- Analyzing the impact of feature persistence information on performance.
Main Results:
- RL methods without 3D maps can support geometrically consistent spatial tasks.
- A designed, geometrically consistent representation improved performance.
- Information about feature persistence (e.g., distant features) enhanced geometric task performance.
- Demonstrated effective non-Cartesian spatial representations.
Conclusions:
- Reinforcement learning provides a viable framework for studying spatial representations.
- Non-Cartesian, learned representations are promising alternatives to cognitive maps.
- These findings stimulate research into new models for spatial perception and navigation.
Related Concept Videos
Observational Learning
717
Albert Bandura's observational learning, also known as imitation or modeling, occurs when a person observes and imitates another's behavior. It is a quicker process than operant conditioning. A well-known example is the Bobo doll study, where children who saw an adult acting aggressively towards the doll were more likely to act aggressively when left alone, compared to those who observed a nonaggressive adult. Many psychologists view observational learning as a form of latent learning...
717
State Space Representation
441
The frequency-domain technique, commonly used in analyzing and designing feedback control systems, is effective for linear, time-invariant systems. However, it falls short when dealing with nonlinear, time-varying, and multiple-input multiple-output systems. The time-domain or state-space approach addresses these limitations by utilizing state variables to construct simultaneous, first-order differential equations, known as state equations, for an nth-order system.
Consider an RLC circuit, a...
Consider an RLC circuit, a...
441
Cognitive Learning
902
Cognitive learning is based on purposive behavior, incidental learning, and insight learning.
E. C. Tolman's theory of purposive behavior emphasizes that much behavior is goal-directed. He argued that to understand behavior, we must look at the entire sequence of actions leading to a goal. For instance, high school students study hard, not just due to past reinforcement but also to achieve the goal of getting into a good college.
Tolman introduced the idea that behavior is influenced by...
E. C. Tolman's theory of purposive behavior emphasizes that much behavior is goal-directed. He argued that to understand behavior, we must look at the entire sequence of actions leading to a goal. For instance, high school students study hard, not just due to past reinforcement but also to achieve the goal of getting into a good college.
Tolman introduced the idea that behavior is influenced by...
902
Collisions in Multiple Dimensions: Problem Solving
5.0K
In multiple dimensions, the conservation of momentum applies in each direction independently. Hence, to solve collisions in multiple dimensions, we should write down the momentum conservation in each direction separately. To help understand collisions in multiple dimensions, consider an example.
A small car of mass 1,200 kg traveling east at 60 km/h collides at an intersection with a truck of mass 3,000 kg traveling due north at 40 km/h. The two vehicles are locked together. What is the...
A small car of mass 1,200 kg traveling east at 60 km/h collides at an intersection with a truck of mass 3,000 kg traveling due north at 40 km/h. The two vehicles are locked together. What is the...
5.0K
Associative Learning
1.0K
Associative learning is a fundamental concept in behavioral psychology, wherein a connection is established between two stimuli or events, leading to a learned response. This process is critical in understanding how behaviors are acquired and modified. Conditioning, the mechanism through which associations are formed, can be divided into two main types: classical conditioning and operant conditioning, each elucidating different aspects of associative learning.
Classical conditioning, also known...
Classical conditioning, also known...
1.0K

