Related Experiment Video
Updated: May 21, 2025

An Open-Source, Fully Customizable 5-Choice Serial Reaction Time Task Toolbox for Automated Behavioral Training of Rodents
Published on: January 19, 2022
Comparative Analysis of Reinforcement Learning Algorithms for Finding Reaction Pathways: Insights from a Large
Yoshihiro Matsumura1, Koji Tabata1,2,3, Tamiki Komatsuzaki1,2,4,5,6
1Institute for Chemical Reaction Design and Discovery (ICReDD), Hokkaido University, Sapporo 001-0020, Japan.
Abstract:
The identification of kinetically feasible reaction pathways that connect a reactant to its product, including numerous intermediates and transition states, is crucial for predicting chemical reactions and elucidating reaction mechanisms. However, as molecular systems become increasingly complex or larger, the number of local minimum structures and transition states grows, which makes this task challenging, even with advanced computational approaches. We introduced a reinforcement learning algorithm to efficiently identify a kinetically feasible reaction pathway between a given local minimum structure for the reactant and a given one for the product, starting from the reactant. The performance of the algorithm was validated using a benchmark data set of large-scale chemical reaction path networks. Several search policies were proposed, using metrics based on energetic or structural similarity to the product's goal structure, for each local minimum structure candidate found during the search. The performances of baseline greedy, random, and uniform search policies varied substantially depending on the system. In contrast, exploration-exploitation balanced policies such as Thompson sampling, probability of improvement, and expected improvement consistently demonstrated stable and high performance. Furthermore, we characterized the search mechanisms that depend on different policies in detail. This study also addressed potential avenues for further research, such as hierarchical reinforcement learning and multiobjective optimization, which could deepen the problem setting explored in this study.
More Related Videos
05:41A Step-by-Step Implementation of DeepBehavior, Deep Learning Toolbox for Automated Behavior Analysis
Published on: February 6, 2020
07:42An Automated T-maze Based Apparatus and Protocol for Analyzing Delay- and Effort-based Decision Making in Free Moving Rodents
Published on: August 2, 2018
Related Concept Videos
Predicting Reaction Outcomes
Reaction Quotient