Related Experiment Video
Updated: Apr 15, 2026

A Machine Learning Approach to Design an Efficient Selective Screening of Mild Cognitive Impairment
Published on: January 11, 2020
Predicting the EQ-5D-3L Preference Index from the SF-12 Health Survey in a National US Sample: A Finite Mixture
Marcelo Coca Perraillon1, Ya-Chen Tina Shih2, Ronald A Thisted3
1Department of Public Health Sciences, University of Chicago, Chicago, IL (MCP)
Background:
. When data on preferences are not available, analysts rely on condition-specific or generic measures of health status like the SF-12 for predicting or mapping preferences. Such prediction is challenging because of the characteristics of preference data, which are bounded, have multiple modes, and have a large proportion of observations clustered at values of 1.
Methods:
. We developed a finite mixture model for cross-sectional data that maps the SF-12 to the EQ-5D-3L preference index. Our model characterizes the observed EQ-5D-3L index as a mixture of 3 distributions: a degenerate distribution with mass at values indicating perfect health and 2 censored (Tobit) normal distributions. Using estimation and validation samples derived from the Medical Expenditure Panel Survey 2000 dataset, we compared the prediction performance of these mixture models to that of 2 previously proposed methods: ordinary least squares regression (OLS) and two-part models.
Results:
. Finite mixture models in which predictions are based on classification outperform two-part models and OLS regression based on mean absolute error, with substantial improvement for samples with fewer respondents in good health. The potential for misclassification is reflected on larger root mean square errors. Moreover, mixture models underperform around the center of the observed distribution.
Conclusions:
. Finite mixtures offer a flexible modeling approach that can take into account idiosyncratic characteristics of the distribution of preferences. The use of mixture models allows researchers to obtain estimates of health utilities when only summary scores from the SF-12 and a limited number of demographic characteristics are available. Mixture models are particularly useful when the target sample does not have a large proportion of individuals in good health.
More Related Videos
13:54A Workflow for Lipid Nanoparticle LNP Formulation Optimization using Designed Mixture-Process Experiments and Self-Validated Ensemble Models SVEM
Published on: August 18, 2023
06:55Inverse Probability of Treatment Weighting Propensity Score using the Military Health System Data Repository and National Death Index
Published on: January 8, 2020
Related Concept Videos
Expected Frequencies in Goodness-of-Fit Tests
Mechanistic Models: Compartment Models in Individual and Population Analysis
Pharmacodynamic Models: Additive and Proportional Drug Effect Model
Distributions to Estimate Population Parameter
Prediction Intervals
However, the point estimate is most likely not the exact value of the population parameter, but close to it. After calculating point estimates, we construct interval estimates, called confidence intervals or prediction intervals. This prediction interval comprises a range of values unlike the point estimate and is a better predictor of the observed sample value, y.
Methods of Medium Optimization