Related Experiment Video
Updated: Jul 5, 2025

O-cresol Concentration Online Measurement Based On Near Infrared Spectroscopy Via Partial Least Square Regression
Published on: November 8, 2019
An empirical comparison of some missing data treatments in PLS-SEM
Lateef Babatunde Amusa1,2, Twinomurinzi Hossana1
1Centre for Applied Data Science, College of Business and Economics, University of Johannesburg, Johannesburg, South Africa.
None:
PLS-SEM is frequently used in applied studies as an excellent tool for examining causal-predictive associations of models for theory development and testing. Missing data are a common problem in empirical analysis, and PLS-SEM is no exception. A comprehensive review of the PLS-SEM literature reveals a high preference for the listwise deletion and mean imputation methods in dealing with missing values. PLS-SEM researchers often disregard strategies for addressing missing data, such as regression imputation and imputation based on the Expectation Maximization (EM) algorithm. In this study, we investigate the utility of these underutilized techniques for dealing with missing values in PLS-SEM and compare them with mean imputation and listwise deletion. Monte Carlo simulations were conducted based on two prominent social science models: the European Customer Satisfaction Index (ECSI) and the Unified Theory of Acceptance and Use of Technology (UTAUT). Our simulation experiments reveal the outperformance of the regression imputation against the other alternatives in the recovery of model parameters and precision of parameter estimates. Hence, regression imputation merit more widespread adoption for treating missing values when analyzing PLS-SEM studies.
Related Concept Videos
Comparing the Survival Analysis of Two or More Groups
Mechanistic Models: Compartment Models in Individual and Population Analysis
Mechanistic Models: Compartment Models in Algorithms for Numerical Problem Solving
In individual population analyses, different algorithms are employed, such as Cauchy's method, which uses a...
Statistical Inference Techniques in Hypothesis Testing: Parametric Versus Nonparametric Data
Parametric statistics, as the name suggests, assumes that data follow a specific distribution, often a normal distribution. This assumption enables robust hypothesis testing and estimation. Parametric methods, like the Student's t-test or Goodness-of-fit test, are frequently employed in biostatistics due to their robustness. For instance,...
Truncation in Survival Analysis
Left truncation occurs when individuals who experienced the event of interest before a certain time are not included in the study. This is often due to a "delayed entry" into the study where only those who survive until a certain entry point are...
Parametric Survival Analysis: Weibull and Exponential Methods
Weibull Distribution
The Weibull distribution is a flexible model used in parametric survival analysis. It can handle both increasing and decreasing hazard rates, depending on its shape parameter...

