基于树的多目标强化学习用于估计耐受性的动态治疗方案

Yao Song1, Lu Wang1

  • 1Department of Biostatistics, University of Michigan, Ann Arbor, MI 48105, United States.

Biometrics
|February 16, 2024
PubMed
概括

这项研究引入了个性化医疗的"宽容方案"概念,在优先事项冲突时提供多种可行的治疗规则. 新的多目标基于树的强化学习 (MOT-RL) 方法有效地估计了这些耐受性动态治疗方案 (tDTR).

相关概念视频

Survival Tree01:19

Survival Tree

Survival trees are a non-parametric method used in survival analysis to model the relationship between a set of covariates and the time until an event of interest occurs, often referred to as the "time-to-event" or "survival time." This method is particularly useful when dealing with censored data, where the event has not occurred for some individuals by the end of the study period, or when the exact time of the event is unknown.
 Building a Survival Tree
Constructing a...
85
Operant Conditioning Intervention01:24

Operant Conditioning Intervention

Operant conditioning serves as a foundational principle in therapeutic interventions aimed at modifying maladaptive behaviors. Central to this approach is the notion that behaviors, both adaptive and maladaptive, are learned through reinforcement. By analyzing the environmental factors that reinforce problematic behaviors, clinicians can design interventions to weaken these reinforcements and replace maladaptive behaviors with healthier alternatives.
In operant conditioning, behaviors that are...
56
Reinforcement Schedules01:24

Reinforcement Schedules

Positive reinforcement is a powerful method for teaching new behaviors to both animals and humans. B.F. Skinner demonstrated this with his experiments using rats in a Skinner box. When a rat pressed a lever, it received a food pellet. This immediate reward encouraged the rat to repeat the behavior. This method, where a reward follows every instance of the behavior, is known as continuous reinforcement. It is highly effective for establishing new behaviors quickly.
Once a behavior is learned,...
147