Related Experiment Video
Updated: Jan 20, 2026

Novel Apparatus and Method for Drug Reinforcement
Published on: August 20, 2010
Reinforcement learning in clinical medicine: a method to optimize dynamic treatment regime over time
1Department of Emergency Medicine, Sir Run Run Shaw Hospital, Zhejiang University School of Medicine, Hangzhou 310016, China.
Abstract:
Precision medicine requires individualized treatment regime for subjects with different clinical characteristics. Machine learning methods have witnessed rapid progress in recent years, which can be employed to make individualized treatment regime in clinical practice. The idea of reinforcement learning method is to take action in response to the changing environment. In clinical medicine, this idea can be used to assign optimal regime to patients with distinct characteristics. In the field of statistics, reinforcement learning has been widely investigated, aiming to identify an optimal dynamic treatment regime (DTR). Q-learning is among the earliest methods to identify optimal DTR, which fits linear outcome models in a recursive manner. The advantage is its easy interpretation and can be performed in most statistical software. However, it suffers from the risk of misspecification of the linear model. More recently, some other methods not so heavily depend on model specification have been developed such as inverse probability weighted estimator and augmented inverse probability weighted estimator. This review introduces the basic ideas of these methods and shows how to perform the learning algorithm within R environment.
More Related Videos
04:09Predicting Treatment Response to Image-Guided Therapies Using Machine Learning: An Example for Trans-Arterial Treatment of Hepatocellular Carcinoma
Published on: October 10, 2018
08:29Ex Vivo Treatment Response of Primary Tumors and/or Associated Metastases for Preclinical and Clinical Development of Therapeutics
Published on: October 2, 2014
Related Concept Videos
07:32Novel Apparatus and Method for Drug Reinforcement
Positive Reinforcement Studies
04:09Predicting Treatment Response to Image-Guided Therapies Using Machine Learning: An Example for Trans-Arterial Treatment of Hepatocellular Carcinoma
08:29Ex Vivo Treatment Response of Primary Tumors and/or Associated Metastases for Preclinical and Clinical Development of Therapeutics
09:32Time-Resolved, Dynamic Computed Tomography Angiography for Characterization of Aortic Endoleaks and Treatment Guidance via 2D-3D Fusion-Imaging
06:14Optimized LC-MS/MS Method for the High-throughput Analysis of Clinical Samples of Ivacaftor, Its Major Metabolites, and Lumacaftor in Biological Fluids of Cystic Fibrosis Patients