Related Experiment Video
Updated: May 5, 2026

Inverse Probability of Treatment Weighting Propensity Score using the Military Health System Data Repository and National Death Index
Published on: January 8, 2020
Robust estimation of optimal dynamic treatment regimes for sequential treatment decisions
Baqun Zhang1, Anastasios A Tsiatis, Eric B Laber
1Department of Preventive Medicine, 680 N. Lakeshore Drive, Suite 1400 Northwestern University, Chicago, Illinois, 60611 U.S.A.
Abstract:
A dynamic treatment regime is a list of sequential decision rules for assigning treatment based on a patient's history. Q- and A-learning are two main approaches for estimating the optimal regime, i.e., that yielding the most beneficial outcome in the patient population, using data from a clinical trial or observational study. Q-learning requires postulated regression models for the outcome, while A-learning involves models for that part of the outcome regression representing treatment contrasts and for treatment assignment. We propose an alternative to Q- and A-learning that maximizes a doubly robust augmented inverse probability weighted estimator for population mean outcome over a restricted class of regimes. Simulations demonstrate the method's performance and robustness to model misspecification, which is a key concern.
More Related Videos
Related Concept Videos
Dosage Regimens: Designs and Approaches
Determination of Multiple Dosing Parameters: Steady-State, Minimum and Maximum Concentrations
Determination of Multiple Dosing Parameters: Loading and Maintenance Doses
Dosage Regimens: Partial Pharmacokinetic Parameters
Dosage Regimen Designs: Nomograms and Tabulations
Pharmacokinetic–Pharmacodynamic Relationship: Problems

