Related Experiment Video
Updated: Jun 3, 2026

'Boden Food Plate': Novel Interactive Web-based Method for the Assessment of Dietary Intake
Published on: September 18, 2018
Exploring statistical approaches to diminish subjectivity of cluster analysis to derive dietary patterns: The
Geraldine Lo Siou1, Yutaka Yasui, Ilona Csizmadi
1Department of Population Health Research, Alberta Health Services—Cancer Care, c/o Holy Cross Site, Box ACB, 2210 2nd Street SW, Calgary, Alberta, Canada T2S 3C3. geraldine.losiou@albertahealthservices.ca
Abstract:
Dietary patterns derived by cluster analysis are commonly reported with little information describing how decisions are made at each step of the analytical process. Using food frequency questionnaire data obtained in 2001-2007 on Albertan men (n = 6,445) and women (n = 10,299) aged 35-69 years, the authors explored the use of statistical approaches to diminish the subjectivity inherent in cluster analysis. Reproducibility of cluster solutions, defined as agreement between 2 cluster assignments, by 3 clustering methods (Ward's minimum variance, flexible beta, K means) was evaluated. Ratios of between- versus within-cluster variances were examined, and health-related variables across clusters in the final solution were described. K means produced cluster solutions with the highest reproducibility. For men, 4 clusters were chosen on the basis of ratios of between- versus within-cluster variances, but for women, 3 clusters were chosen on the basis of interpretability of cluster labels and descriptive statistics. In comparison with those in other clusters, men and women in the "healthy" clusters by greater proportions reported normal body mass index, smaller waist circumference, and lower energy intakes. The authors' approach appeared helpful when choosing the clustering method for both sexes and the optimal number of clusters for men, but additional analyses are required to understand why it performed differently for women.
Related Concept Videos
Statistical Methods for Analyzing Epidemiological Data
Sampling Plans
Random sampling is a method where each member of the population has an equal chance of being selected for the sample. It involves selecting individuals randomly, often using random number generators or lottery-type methods. For example, when analyzing the properties of a...
Regression Toward the Mean
Cluster Sampling Method
To choose a cluster sample, divide the population into clusters (groups) and then randomly select some of the clusters. All the members from these clusters are in the cluster sample. For example, if you randomly sample four departments from your...
Study Design in Statistics
Does aspirin reduce the risk of heart attacks? Is one brand of fertilizer more effective at growing roses than another? Is fatigue as dangerous to a driver as the influence of alcohol? Questions like these are answered using randomized experiments with proper...
Model-Independent Approaches for Pharmacokinetic Data: Noncompartmental Analysis
One important characteristic of noncompartmental analyses is that drug exposure increases proportionally with increasing doses. This relationship...