Related Experiment Video
Updated: Sep 9, 2025

Methodology for Establishing a Community-Wide Life Laboratory for Capturing Unobtrusive and Continuous Remote Activity and Health Data
Published on: July 27, 2018
A Robust Mixed-Effects Bandit Algorithm for Assessing Mobile Health Interventions
Easton K Huch1, Jieru Shi2, Madeline R Abbott2
1Department of Statistics, University of Michigan, Ann Arbor, MI 48109, USA.
None:
Mobile health leverages personalized, contextually-tailored interventions optimized through bandit and reinforcement learning algorithms. Despite its promise, challenges like participant heterogeneity, nonstationarity, and nonlinearity in rewards hinder algorithm performance. We propose a robust contextual bandit algorithm, termed "DML-TS-NNR", that simultaneously addresses these challenges via (1) modeling the differential reward with user- and time-specific incidental parameters, (2) network cohesion penalties, and (3) debiased machine learning for flexible estimation of baseline rewards. We establish a high-probability regret bound that depends solely on the dimension of the differential reward model. This feature enables us to achieve robust regret bounds even when the baseline reward is highly complex. We demonstrate the superior performance of the DML-TS-NNR algorithm in a simulation and two off-policy evaluation studies.
Related Concept Videos
Randomized Experiments
Simple randomization
Simple...
Mechanistic Models: Compartment Models in Individual and Population Analysis

