Related Experiment Video
Updated: May 21, 2025

Applying an eMASS Customization Program as a Research Tool to Evaluate Consumer Benefits
Published on: September 27, 2019
Bayesian nonparametric latent class analysis with different item types
Meng Qiu1, Sally Paganin2, Ilsang Ohn3
1Department of Psychological Sciences, University of California, Merced.
Abstract:
Latent class analysis (LCA) requires deciding on the number of classes. This is traditionally addressed by fitting several models with an increasing number of classes and determining the optimal one using model selection criteria. However, different criteria can suggest different models, making it difficult to reach a consensus on the best criterion. Bayesian nonparametric LCA based on the Dirichlet process mixture (DPM) model is a flexible alternative approach that allows for inferring the number of classes from the data. In this article, we introduce a DPM-based mixed-mode LCA model, referred to as DPM-MMLCA, which clusters individuals based on indicators measured on mixed metrics. We illustrate two algorithms for posterior estimation and discuss inferential procedures to estimate the number of classes and their composition. A simulation study is conducted to compare the performance of the DPM-MMLCA with the traditional mixed-mode LCA under different scenarios. Five design factors are considered, including the number of latent classes, the number of observed variables, sample size, mixing proportions, and class separation. Performance measures include evaluating the correct identification of the number of latent classes, parameter recovery, and assignment of class labels. The Bayesian nonparametric LCA approach is illustrated using three real data examples. Additionally, a hands-on tutorial using R and the nimble package is provided for ease of implementation. (PsycInfo Database Record (c) 2025 APA, all rights reserved).
Related Concept Videos
Statistical Inference Techniques in Hypothesis Testing: Parametric Versus Nonparametric Data
Parametric statistics, as the name suggests, assumes that data follow a specific distribution, often a normal distribution. This assumption enables robust hypothesis testing and estimation. Parametric methods, like the Student's t-test or Goodness-of-fit test, are frequently employed in biostatistics due to their robustness. For instance,...
How Data are Classified: Categorical Data
Data are classified based on whether they are measurable or not. Categorical data cannot be measured; instead, it can be divided into categories. For example, if Y denotes a person's party affiliation, some examples of Y include...
Law of Independent Assortment
Regression Analysis
In regression analysis, a regression equation is determined based on the line of best fit– a line that best fits the data points plotted in a graph. This line is also called the regression line. The algebraic equation for the regression line is called the regression equation. It is represented as:
Multiple Regression
Farmers can use multiple regression to determine the crop yield based on more than one factor, such as water availability, fertilizer, soil properties, etc. Here, the crop yield is the response or dependent variable as it depends on the other independent variables. The analysis requires the construction of a scatter plot...
Goodness-of-Fit Test

