Related Experiment Video
Updated: Apr 24, 2026

Comprehensive Profiling of Dopamine Regulation in Substantia Nigra and Ventral Tegmental Area
Published on: August 10, 2012
Multi-neurotransmitter synergistically regulated basal ganglia reinforcement learning model
Si-Ping Zhang1, Wei-Guo Jiang2, Xu Deng2
1The Key Laboratory of Biomedical Information Engineering of Ministry of Education, Institute of Health and Rehabilitation Science, School of Life Science and Technology, Xi'an Jiaotong University, Xi'an, Shaanxi, China; Research Center for Brain-Inspired Intelligence, Xi'an Jiaotong University, Xi'an, Shaanxi, China; State Key Laboratory of Human-Machine Hybrid Augmented Intelligence, Xi'an Jiaotong University, Xi'an, Shaanxi, China.
None:
As a core brain region for motor control and cognitive decision-making, the basal ganglia relies on the dynamic balance among direct, indirect, and hyperdirect pathways, which is maintained by the synergistic action of multiple neurotransmitters. To address the critical limitations of traditional reinforcement learning algorithms, including low exploration efficiency and poor environmental adaptability, this study proposes a multi-neurotransmitter-regulated basal ganglia reinforcement learning (MNBG-RL) model inspired by the neurocomputational principles of the cortex-basal ganglia-thalamus loop. The model features three core innovations: First, a dual neurotransmitter regulation system of dopamine (DA) and noradrenaline (NA) is constructed, where DA encodes reward prediction errors and dynamically modulates pathway weights through D1/D2 receptors, while NA, under the regulation of DA, enables precise generation of exploration intensity by modulating subthalamic nucleus (STN) states. Second, a two-dimensional Ising network is embedded in the indirect pathway to simulate the population activity of STN neurons, generating adaptive structured exploration signals (replacing traditional blind Gaussian noise) that support smooth state transitions among ordered, critical, and disordered regimes. Third, the model performance is validated on three typical tasks: (1) the model flexibly adapts to environmental reversals and converges to optimal policies in complex decision-making tasks; (2) clamping DA/NA levels (simulating impaired neurotransmitter systems) significantly reduces decision flexibility and even leads to task failure; (3) compared with traditional exploration methods, the MNBG-RL model achieves substantially improved convergence rates and greatly reduced policy oscillation. This study not only supports the development of biologically plausible basal ganglia-inspired reinforcement learning frameworks but also provides a computational platform for modeling decision-making deficits in neurological disorders such as Parkinson's disease.
More Related Videos
Related Concept Videos
Role of Neurotransmitters in Memory
Glutamate and Synaptic Plasticity
Glutamate, the brain's main excitatory neurotransmitter, is...
Neural Regulation
Neurochemical Transmission: Sites of Drug Action
Long-term Potentiation
Long-term Potentiation
Hebbian LTP
LTP can occur when...
Classification of Neurotransmitters

