在计数回归模型中隐性变量之间的相互作用
Christoph Kiefer1, Sarah Wilker2, Axel Mayer3
1Methods and Evaluation, Department of Psychology, Bielefeld University, Universitätsstraße 25, D-33501, Bielefeld, Germany. christoph.kiefer@uni-bielefeld.de.
Behavior research methods
|August 26, 2024
概括
研究人员经常忽视计数回归模型中的测量误差,导致结果偏差. 一个新的隐性变量计数回归模型 (LV-CRM) 准确地估计了系数,并改善了统计推理,即使有隐性相互作用.
科学领域:
- 心理学和社会科学 心理学和社会科学
- 统计建模 统计建模
背景情况:
- 计算结果变量在心理学和社会科学中很常见.
- 计数数据的通用线性模型 (GLM) 通常忽略预测器中的测量误差,导致减弱偏差.
- 现有的方法很少解决在计数回归中涉及潜在变量的相互作用.
研究的目的:
- 引入一个隐性变量计数回归模型 (LV-CRM),该模型包含隐性预测因素及其相互作用.
- 与基于GLM的计数回归模型相比,评估LV-CRM的估计准确性和统计推断.
- 展示LV-CRM在临床心理学中的实际应用.
主要方法:
- 开发了一个潜在变量计数回归模型 (LV-CRM).
- 进行了三项模拟研究,将LV-CRM与基于GLM的计数回归模型进行比较.
- 在各种条件下调查估计准确性和统计推断.
主要成果:
- 基于GLM的模型显示了回归系数的严重偏差,即使具有高预测器可靠性.
- LV-CRM提供了几乎无偏见的回归系数,即使样本大小适度.
- 对于LV-CRM来说,统计推断通常是可以接受的,而基于GLM的模型显示出混合的结果 (覆盖率低,可接受的检测率).
结论:
- LV-CRM有效地考虑了隐性预测器中的测量误差及其在计数回归中的相互作用.
- 基于GLM的传统方法,LV-CRM为数量数据分析提供了更准确,更可靠的替代方案.
- 拟议的框架对于处理复杂计数数据结构的心理学和社会科学研究人员来说是有价值的.
相关概念视频
Multiple Regression
3.0K
Multiple regression assesses a linear relationship between one response or dependent variable and two or more independent variables. It has many practical applications.
Farmers can use multiple regression to determine the crop yield based on more than one factor, such as water availability, fertilizer, soil properties, etc. Here, the crop yield is the response or dependent variable as it depends on the other independent variables. The analysis requires the construction of a scatter plot...
Farmers can use multiple regression to determine the crop yield based on more than one factor, such as water availability, fertilizer, soil properties, etc. Here, the crop yield is the response or dependent variable as it depends on the other independent variables. The analysis requires the construction of a scatter plot...
3.0K
Mechanistic Models: Compartment Models in Individual and Population Analysis
33
Mechanistic models are utilized in individual analysis using single-source data, but imperfections arise due to data collection errors, preventing perfect prediction of observed data. The mathematical equation involves known values (Xi), observed concentrations (Ci), measurement errors (εi), model parameters (ϕj), and the related function (ƒi) for i number of values. Different least-squares metrics quantify differences between predicted and observed values. The ordinary least...
33
Friedman Two-way Analysis of Variance by Ranks
170
Friedman's Two-Way Analysis of Variance by Ranks is a nonparametric test designed to identify differences across multiple test attempts when traditional assumptions of normality and equal variances do not apply. Unlike conventional ANOVA, which requires normally distributed data with equal variances, Friedman's test is ideal for ordinal or non-normally distributed data, making it particularly useful for analyzing dependent samples, such as matched subjects over time or repeated measures...
170
Correlation and Regression
1.2K
In statistics, correlation describes the degree of association between two variables. In the subfield of linear regression, correlation is mathematically expressed by the correlation coefficient, which describes the strength and direction of the relationship between two variables. The coefficient is symbolically represented by 'r' and ranges from -1 to +1. A positive value indicates a positive correlation where the two variables move in the same direction. A negative value suggests a...
1.2K
Regression Analysis
5.7K
Regression analysis is a statistical tool that describes a mathematical relationship between a dependent variable and one or more independent variables.
In regression analysis, a regression equation is determined based on the line of best fit– a line that best fits the data points plotted in a graph. This line is also called the regression line. The algebraic equation for the regression line is called the regression equation. It is represented as:
In regression analysis, a regression equation is determined based on the line of best fit– a line that best fits the data points plotted in a graph. This line is also called the regression line. The algebraic equation for the regression line is called the regression equation. It is represented as:
5.7K
Biostatistics: Overview
227
Biostatistics plays a crucial role in understanding and analyzing data in healthcare and biology. Biostatisticians conduct experiments, gather evidence, and draw meaningful conclusions using statistical methods and techniques. Different variables form the foundation of biostatistical analysis, allowing researchers to understand and interpret data effectively. These variables are classified into different types, each serving a specific purpose in statistical analysis.
Discrete variables are...
Discrete variables are...
227


