Related Experiment Video
Updated: Oct 15, 2025

Assessing Changes in Volatile General Anesthetic Sensitivity of Mice after Local or Systemic Pharmacological Intervention
Published on: October 16, 2013
Patient-Specific Sedation Management via Deep Reinforcement Learning
Niloufar Eghbali1, Tuka Alhanai2, Mohammad M Ghassemi1
1Human Augmentation and Artificial Intelligence Laboratory, Department of Computer Science, Michigan State University, East Lansing, MI, United States.
Abstract:
Introduction: Developing reliable medication dosing guidelines is challenging because individual dose-response relationships are mitigated by both static (e. g., demographic) and dynamic factors (e.g., kidney function). In recent years, several data-driven medication dosing models have been proposed for sedatives, but these approaches have been limited in their ability to assess interindividual differences and compute individualized doses. Objective: The primary objective of this study is to develop an individualized framework for sedative-hypnotics dosing. Method: Using publicly available data (1,757 patients) from the MIMIC IV intensive care unit database, we developed a sedation management agent using deep reinforcement learning. More specifically, we modeled the sedative dosing problem as a Markov Decision Process and developed an RL agent based on a deep deterministic policy gradient approach with a prioritized experience replay buffer to find the optimal policy. We assessed our method's ability to jointly learn an optimal personalized policy for propofol and fentanyl, which are among commonly prescribed sedative-hypnotics for intensive care unit sedation. We compared our model's medication performance against the recorded behavior of clinicians on unseen data. Results: Experimental results demonstrate that our proposed model would assist clinicians in making the right decision based on patients' evolving clinical phenotype. The RL agent was 8% better at managing sedation and 26% better at managing mean arterial compared to the clinicians' policy; a two-sample t-test validated that these performance improvements were statistically significant (p < 0.05). Conclusion: The results validate that our model had better performance in maintaining control variables within their target range, thereby jointly maintaining patients' health conditions and managing their sedation.
More Related Videos
05:39Author Spotlight: A Non-Intubated Video-Assisted Thoracoscopic Surgery with Multimodal Analgesia and Sevoflurane Inhalation Anesthesia
Published on: May 26, 2023
09:22A Method for Remotely Silencing Neural Activity in Rodents During Discrete Phases of Learning
Published on: June 22, 2015
Related Concept Videos
Stages of General Anesthesia
Parenteral Anesthetics: Overview
Skeletal Muscle Relaxants: Therapeutic Uses
Sedatives and Hypnotics: Overview
Sedative-hypnotics are categorized into barbiturates, benzodiazepines (BZDs), and non-benzodiazepines or Z-drugs. These drugs work by suppressing central nervous system activity, and this suppression is dose-dependent. Older sedative medications, like barbiturates, follow a linear curve in...
Skeletal Muscle Relaxants: Adverse Effects
Unlike...
General Anesthesia: Overview
General anesthesia induces unconsciousness in the whole body, while the others target specific areas or sensations. It is administered to minimize adverse effects, maintain...