Related Experiment Video
Updated: Jan 8, 2026

08:07
Assembly and Characterization of Biomolecular Memristors Consisting of Ion Channel-doped Lipid Membranes
Published on: March 9, 2019
8.3K
Actor-critic networks with analogue memristors mimicking reward-based learning
Kevin Portner1, Till Zellweger1, Flavio Martinelli2
1Integrated Systems Laboratory, ETH Zurich, Zurich, Switzerland.
Nature Machine Intelligence
|December 24, 2025
Summary
Researchers developed a novel bio-inspired computing system using analogue memristors for reinforcement learning. This hardware performs online training and action selection entirely in memory, advancing neuromorphic computing engines.
Area of Science:
- Neuromorphic Engineering
- Computational Neuroscience
Background:
- Current bio-inspired computing often mimics only parts of the brain.
- Memristive hardware is typically used for limited functions in learning algorithms, requiring software for complex tasks.
Purpose of the Study:
- To demonstrate a fully integrated reinforcement learning system using analogue memristors.
- To create a bio-inspired neural network architecture that performs reward-based learning entirely in hardware.
Main Methods:
- Implemented an actor-critic temporal difference algorithm on analogue memristors.
- Utilized memristors as multi-purpose elements: synaptic weights, weight update calculators, and action determiners.
- Tested the framework on T-maze and Morris water maze navigation tasks.
Main Results:
- Achieved online weight training and direct hardware calculation of temporal difference errors.
- Enabled complete in-memory computation, eliminating data movement for weight training.
- Successfully demonstrated navigation in simulated environments using the memristor-based system.
Conclusions:
- This work presents the first fully in-memory, online reinforcement learning system for bio-inspired computing.
- The approach paves the way for more efficient and brain-like neuromorphic computing engines.
- Analogue memristors can serve as versatile components for advanced artificial intelligence hardware.
Related Concept Videos
Observational Learning
782
Albert Bandura's observational learning, also known as imitation or modeling, occurs when a person observes and imitates another's behavior. It is a quicker process than operant conditioning. A well-known example is the Bobo doll study, where children who saw an adult acting aggressively towards the doll were more likely to act aggressively when left alone, compared to those who observed a nonaggressive adult. Many psychologists view observational learning as a form of latent learning...
782
Design Example: Frog Muscle Response
544
A student is tasked to work on an intriguing experiment involving an RL (Resistor-Inductor) circuit to study the muscle response of a frog's leg to electrical stimulation. The RL circuit plays a crucial role in this experiment, providing the means to control and measure the electrical impulses that trigger muscle contraction.
When the switch connecting the RL circuit is closed, a brief muscle contraction is observed. This is because, at a steady state, the inductor acts like a short...
When the switch connecting the RL circuit is closed, a brief muscle contraction is observed. This is because, at a steady state, the inductor acts like a short...
544

