Related Experiment Video
Updated: Jan 19, 2026

Stereoacuity Improvement using Random-Dot Video Games
Published on: January 14, 2020
Improved robustness of reinforcement learning policies upon conversion to spiking neuronal network platforms applied
Devdhar Patel1, Hananel Hazan1, Daniel J Saunders1
1Biologically Inspired Neural and Dynamical Systems Laboratory (BINDS) College of Computer and Information Sciences, 140 Governors Drive, University of Massachusetts Amherst, Amherst, MA 01003, USA.
Abstract:
Deep Reinforcement Learning (RL) demonstrates excellent performance on tasks that can be solved by trained policy. It plays a dominant role among cutting-edge machine learning approaches using multi-layer Neural networks (NNs). At the same time, Deep RL suffers from high sensitivity to noisy, incomplete, and misleading input data. Following biological intuition, we involve Spiking Neural Networks (SNNs) to address some deficiencies of deep RL solutions. Previous studies in image classification domain demonstrated that standard NNs (with ReLU nonlinearity) trained using supervised learning can be converted to SNNs with negligible deterioration in performance. In this paper, we extend those conversion results to the domain of Q-Learning NNs trained using RL. We provide a proof of principle of the conversion of standard NN to SNN. In addition, we show that the SNN has improved robustness to occlusion in the input image. Finally, we introduce results with converting full-scale Deep Q-network to SNN, paving the way for future research to robust Deep RL applications.
Related Concept Videos
06:25Stereoacuity Improvement using Random-Dot Video Games
08:18WheelCon: A Wheel Control-Based Gaming Platform for Studying Human Sensorimotor Control
06:21Installation Method to Enhance Quality Control for Fiber Reinforced Polymer Spike Anchors
11:32A Flexible Platform for Monitoring Cerebellum-Dependent Sensory Associative Learning
11:20Recording Single Neurons' Action Potentials from Freely Moving Pigeons Across Three Stages of Learning
11:18Closed-loop Neuro-robotic Experiments to Test Computational Properties of Neuronal Networks

