估计MST中MLE分数的条件标准测量错误
Yuanyuan J Stirn1, Won-Chan Lee1
1The University of Iowa, Iowa City, USA.
Educational and psychological measurement
|March 2, 2026
概括
本研究介绍了一种分析方法,用于计算多阶段测试 (MST) 中的条件标准测量误差 (CSEM). 该方法在较长的测试中被证明是可靠的,准确度随着测试长度的增加而增加.
科学领域:
- 心理测量 心理测量 心理测量
- 教育测量教育的测量
- 统计建模 统计建模
背景情况:
- 条件标准测量误差 (CSEM) 对于在适应性测试中准确估计能力至关重要.
- 在多阶段测试 (MST) 中进行CSEM计算的现有方法可能是计算密集的或依赖于模拟.
- 需要有效和准确的分析方法来计算MST中的CSEM.
研究的目的:
- 提出和评估一种基于信息的分析方法,用于在多阶段测试 (MST) 中计算CSEM.
- 将拟议的分析方法的准确性与基于模拟的CSEM估计进行比较.
- 评估测试长度和设计复杂度对拟议方法准确性的影响.
主要方法:
- 开发了一种基于信息的分析方法,利用最大概率估计.
- 使用拟议的分析方法计算CSEM,并将其与基于模拟的估计进行比较.
- 在四个不同的MST设计中评估性能,测试长度各不相同.
主要成果:
- 基于分析和模拟的CSEM显示,随着测试长度的增加,结果趋同.
- 拟议的方法为CSEM在较长的MST中提供了可靠的近似.
- 较短的测试和更复杂的MST设计需要更多的项目来实现可比的准确性.
结论:
- 提出的基于信息的分析方法为MST中CSEM计算提供了可靠的方法,特别是在更长的测试中.
- 该方法的准确性受到测试长度和设计复杂性的影响,建议对较短或更复杂的测试进行调整.
- 这种分析方法为MST中CSEM估计提供了基于模拟的方法的实用替代方案.
相关概念视频
Estimating Population Mean with Unknown Standard Deviation
8.9K
In practice, we rarely know the population standard deviation. In the past, when the sample size was large, this did not present a problem to statisticians. They used the sample standard deviation s as an estimate for σ and proceeded as before to calculate a confidence interval with close enough results. However, statisticians ran into problems when the sample size was small. A small sample size caused inaccuracies in the confidence interval.
William S. Gosset (1876–1937) of the...
William S. Gosset (1876–1937) of the...
8.9K
Estimating Population Mean with Known Standard Deviation
9.8K
To construct a confidence interval for a single unknown population mean μ, where the population standard deviation is known, we need sample mean as an estimate for μ and we need the margin of error. Here, the margin of error (EBM) is called the error bound for a population mean (abbreviated EBM). The sample mean is the point estimate of the unknown population mean μ.
The confidence interval estimate will have the form as follows:
(point estimate - error bound, point estimate +...
The confidence interval estimate will have the form as follows:
(point estimate - error bound, point estimate +...
9.8K
Estimating Population Standard Deviation
3.4K
When the population standard deviation is unknown and the sample size is large, the sample standard deviation s is commonly used as a point estimate of σ. However, it can sometimes under or overestimate the population standard deviation. To overcome this drawback, confidence intervals are determined to estimate population parameters and eliminate any calculation bias accurately. However, this only applies to random samples from normally distributed populations. Knowing the sample mean and...
3.4K
Testing a Claim about Standard Deviation
3.0K
A complete procedure to test a claim about population standard deviation or population variance is explained here.
The hypothesis testing for the claim of population standard deviation (or variance) requires the data and samples to be random and unbiased. The population distribution also must be normal. There is no specific requirement on the sample size as the estimation is based on the chi-square distribution.
As a first step, the hypothesis (null and alternative) concerning the claim about...
The hypothesis testing for the claim of population standard deviation (or variance) requires the data and samples to be random and unbiased. The population distribution also must be normal. There is no specific requirement on the sample size as the estimation is based on the chi-square distribution.
As a first step, the hypothesis (null and alternative) concerning the claim about...
3.0K
Standard Error of the Mean
12.7K
The sampling variability of a statistic is defined as how much the statistic varies from one sample to another. The sampling variability of a statistic is typically measured by measuring its standard error.
The standard error of the mean is an example of a standard error. It is a unique standard deviation known as the standard deviation of the sampling distribution of the mean. The standard error of the mean is a statistic that calculates how correctly a sample distribution represents a...
The standard error of the mean is an example of a standard error. It is a unique standard deviation known as the standard deviation of the sampling distribution of the mean. The standard error of the mean is a statistic that calculates how correctly a sample distribution represents a...
12.7K
Multiple Comparison Tests
4.5K
Multiple comparison test, abbreviated as MCT, is a post hoc analysis generally performed after comparing multiple samples with one or more tests. An MCT will help identify a significantly different sample among multiple samples or a factor among multiple factors.
It would be easy to compare two samples using a significance alpha level of 0.05. In other words, there is only one sample pair to be compared. However, it would be difficult to identify a significantly different sample if the number...
It would be easy to compare two samples using a significance alpha level of 0.05. In other words, there is only one sample pair to be compared. However, it would be difficult to identify a significantly different sample if the number...
4.5K


