Related Experiment Video
Updated: May 20, 2025

Polar Histogram Visualization of Acute Stress Disorder Scale Scores for Comprehensive Clinical Assessment
Published on: December 6, 2024
Using XGBoost and SHAP to explain citizens' differences in policy support for reimposing COVID-19 measures in the
Jose Ignacio Hernandez1,2, Sander van Cranenburgh2, Marijn de Bruin3,4
1Center of Economics for Sustainable Development (CEDES), Faculty of Economics and Government, Universidad San Sebastian, Concepción, Chile.
Abstract:
Several studies examined what drives citizens' support for COVID-19 measures, but no works have addressed how the effects of these drivers are distributed at the individual level. Yet, if significant differences in support are present but not accounted for, policymakers' interpretations could lead to misleading decisions. In this study, we use XGBoost, a supervised machine learning model, combined with SHAP (Shapley Additive eXplanations) to identify the factors associated with differences in policy support for COVID-19 measures and how such differences are distributed across different citizens and measures. We use secondary data from a Participatory Value Evaluation (PVE) experiment, in which 1,888 Dutch citizens answered which COVID-19 measures should be imposed under four risk scenarios. We identified considerable heterogeneity in citizens' support for different COVID-19 measures regarding different age groups, the weight given to citizens' opinions and the perceived risk of getting sick of COVID-19. Data analysis methods employed in previous studies do not reveal such heterogeneity of policy support. Policymakers can use our results to tailor measures further to increase support for specific citizens/measures.
Supplementary Information:
The online version contains supplementary material available at 10.1007/s11135-024-01938-2.
Related Concept Videos
Comparing the Survival Analysis of Two or More Groups
Bias in Epidemiological Studies
Residuals and Least-Squares Property
If the observed data point lies above the line, the residual is positive, and the line underestimates the actual data value for y. If the observed data point lies below the line, the residual is negative, and the line overestimates the actual data value for y.
The process of fitting the best-fit...
Statistical Methods for Analyzing Epidemiological Data
Confounding in Epidemiological Studies
Social Proof

