Related Experiment Video
Updated: Oct 19, 2025

11:18
Closed-loop Neuro-robotic Experiments to Test Computational Properties of Neuronal Networks
Published on: March 2, 2015
10.5K
Deep Reinforcement Learning With Modulated Hebbian Plus Q-Network Architecture.
IEEE Transactions on Neural Networks and Learning Systems
|September 24, 2021
Summary
This study introduces a novel bio-inspired neural network, the modulated Hebbian plus Q-network architecture (MOHQA), to solve challenging partially observable Markov decision process (POMDP) problems. MOHQA effectively handles confounding observations and sparse rewards where traditional reinforcement learning methods fail.
Area of Science:
- Artificial Intelligence
- Computational Neuroscience
- Machine Learning
Background:
- Partially Observable Markov Decision Processes (POMDPs) present challenges for standard reinforcement learning (RL) algorithms, particularly when observations are confounding and rewards are sparse.
- Temporal Difference (TD) error, crucial for many RL algorithms like Deep Q-Network (DQN), becomes unreliable in such scenarios due to difficulties in deriving accurate error signals from observations.
Purpose of the Study:
- To address the limitations of existing RL algorithms in solving confounding POMDPs.
- To introduce a novel bio-inspired neural architecture, the modulated Hebbian plus Q-network architecture (MOHQA), designed to overcome these challenges.
Main Methods:
- The proposed MOHQA architecture integrates a modulated Hebbian network (MOHN) with a Deep Q-Network (DQN).
- The MOHN utilizes bio-inspired neural traces to bridge temporal delays between actions and rewards, compensating for inaccurate TD errors.
- DQN handles low-level feature extraction and control, while MOHN assists in high-level decision-making by associating rewards with past states and actions.
Main Results:
- Simulations demonstrated that MOHQA significantly improved upon standard DQN performance.
- MOHQA outperformed several state-of-the-art RL algorithms, including Advantage Actor-Critic (A2C), Quantile Regression DQN with Long Short-Term Memory (QRDQN + LSTM), REINFORCE, and Aggregated Memory for Reinforcement Learning (AMRL).
- The proposed architecture showed particular efficacy on difficult POMDPs characterized by confounding stimuli and sparse rewards.
Conclusions:
- The MOHQA architecture offers a robust solution for confounding POMDPs, outperforming existing methods.
- Combining Hebbian associative learning with deep reinforcement learning provides synergistic advantages for complex decision-making tasks.
- This bio-inspired approach demonstrates the potential for novel neural architectures to advance reinforcement learning capabilities.
Related Concept Videos
Neural Regulation
40.6K
Digestion begins with a cephalic phase that prepares the digestive system to receive food. When our brain processes visual or olfactory information about food, it triggers impulses in the cranial nerves innervating the salivary glands and stomach to prepare for food.
40.6K
Neural Circuits
1.9K
Neural circuits and neuronal pools are two of the main structures found in the nervous system. Neural circuits are networks of neurons that work together to carry out a specific task or process. They consist of interconnected neurons and glial cells, which provide structural and metabolic support.
Neuronal pools are collections of nerve cells with similar functions and interact through chemical and electrical signals. These pools include both interneurons (the central neural circuit nodes that...
Neuronal pools are collections of nerve cells with similar functions and interact through chemical and electrical signals. These pools include both interneurons (the central neural circuit nodes that...
1.9K
Neural Control of Respiration
3.3K
The neural regulation of respiration is a meticulously coordinated process primarily controlled by the respiratory centers located within the brainstem. These centers, composed of specialized neurons, transmit nerve impulses that control the contraction and relaxation of our respiratory muscles.
Respiratory Centers in the Brainstem
Two primary areas comprise the respiratory center: the medullary respiratory center in the medulla oblongata and the pontine respiratory group in the pons. The...
Respiratory Centers in the Brainstem
Two primary areas comprise the respiratory center: the medullary respiratory center in the medulla oblongata and the pontine respiratory group in the pons. The...
3.3K
Propagation of Action Potentials
7.5K
The propagation of an action potential refers to the process by which a nerve impulse, or "action potential," travels along a neuron.
Neurons (nerve cells) have a resting membrane potential, with a slightly negative charge inside compared to outside. This is maintained by ion channels, such as sodium (Na+) and potassium (K+) channels, which control the flow of ions. When a stimulus, like a touch or a signal from another neuron, triggers the neuron, sodium channels open, allowing sodium ions to...
Neurons (nerve cells) have a resting membrane potential, with a slightly negative charge inside compared to outside. This is maintained by ion channels, such as sodium (Na+) and potassium (K+) channels, which control the flow of ions. When a stimulus, like a touch or a signal from another neuron, triggers the neuron, sodium channels open, allowing sodium ions to...
7.5K
Long-term Potentiation
56.5K
Long-term potentiation, or LTP, is one of the ways by which synaptic plasticity—changes in the strength of chemical synapses—can occur in the brain. LTP is the process of synaptic strengthening that occurs over time between pre- and postsynaptic neuronal connections. The synaptic strengthening of LTP works in opposition to the synaptic weakening of long-term depression (LTD) and together are the main mechanisms that underlie learning and memory.
56.5K
Associative Learning
686
Associative learning is a fundamental concept in behavioral psychology, wherein a connection is established between two stimuli or events, leading to a learned response. This process is critical in understanding how behaviors are acquired and modified. Conditioning, the mechanism through which associations are formed, can be divided into two main types: classical conditioning and operant conditioning, each elucidating different aspects of associative learning.
Classical conditioning, also known...
Classical conditioning, also known...
686
