相关实验视频
Updated: Jun 16, 2025

04:40
Tactile Semiautomatic Passive-Finger Angle Stimulator TSPAS
Published on: July 30, 2020
2.9K
非对称正确的人适合z统计 对于拉什丸模型的z统计
Zhongtian Lin1, Tao Jiang2, Frank Rijmen2
1Financial Industry Regulatory Authority, Washington, USA. lzt713@gmail.com.
Psychometrika
|August 17, 2024
概括
这项研究为拉什试卷模型引入了新的人体适应统计数据,并扩展了对物品响应理论的现有方法. 这些统计数据有效地检测复杂的测试结构中的异常反应.
科学领域:
- 心理测量 心理测量 心理测量
- 教育测量教育的测量
- 统计建模 统计建模
背景情况:
- 确立的人适合的统计数据,如和,仅限于单维或联合多维物件响应理论 (IRT) 模型.
- 现有的方法通常需要对所有潜在特征进行联合估计,这给计算带来了挑战.
研究的目的:
- 提出新的人适合统计学,和,特别是对于拉什测试模型.
- 将人体适应性评估的适用性扩展到混合效应的IRT模型.
- 为拟议的统计提供计算算法.
主要方法:
- 基于边缘化最大概率能力估计器的和统计的开发.
- 扩展了Lord-Wingersky算法的计算效率.
- 模拟研究评估I型错误率和功率.
主要成果:
- 拟议的统计表明I型错误率接近名义水平.
- 在检测异常反应方面显示了令人满意的功率.
- 统计数据减少到已建立的模型和单维模型的模型.
结论:
- 新的统计数据提供了一个可靠的方法,用于在Rasch试卷模型中评估人体适应性.
- 这些方法增强了混合结构测试中响应行为的评估.
- 拟议的统计扩大了可以评估个人适合的IRT模型的范围.
相关概念视频
Wald-Wolfowitz Runs Test II
202
The Wald-Wolfowitz runs test, commonly referred to as the runs test, is a nonparametric test used to assess the randomness of ordered data. The test evaluates the number of runs, which are consecutive sequences of similar elements within the data. If the number of runs is significantly higher or lower than expected, the data is considered non-random, indicating a detectable pattern or structure.
For binary data, runs are identified using symbols such as + and −, or equivalently, 1s and...
For binary data, runs are identified using symbols such as + and −, or equivalently, 1s and...
202
Goodness-of-Fit Test
3.3K
The goodness-of-fit test is a type of hypothesis test which determines whether the data "fits" a particular distribution. For example, one may suspect that some anonymous data may fit a binomial distribution. A chi-square test (meaning the distribution for the hypothesis test is chi-square) can be used to determine if there is a fit. The null and alternative hypotheses may be written in sentences or stated as equations or inequalities. The test statistic for a goodness-of-fit test is given as...
3.3K
Choosing Between z and t Distribution
2.8K
The z and the Student t distribution estimate the population mean using the sample mean and standard deviation. However, to decide which distribution to use for a calculation, one needs to determine the sample size, the nature of the distribution, and whether the population standard deviation is known. If the population standard deviation is known and the population is normally distributed, or if the sample size is greater than 30, the z distribution is preferred. The Student t distribution is...
2.8K
Sign Test for Median of Single Population
109
In general, the sign test serves as a nonparametric method to test hypotheses about the median of a single population when the data does not follow a known distribution. This simplicity makes it particularly useful for small sample sizes or when the assumptions of parametric tests cannot be met. The process begins with identifying a null hypothesis, typically stating that the population median equals a specific value. The alternative hypothesis could be that the median is either not equal to,...
109
Introduction to the Sign Test
753
The sign test is an important tool in nonparametric statistics, offering a straightforward yet effective method for analyzing matched pairs, nominal data, or hypotheses concerning the median of a population. It transforms data points into positive or negative signs, avoiding the need for assumptions about data distribution and instead focusing on the direction of change. It is particularly valuable when data does not conform to the normal distribution requirements of many parametric tests. For...
753
Expected Frequencies in Goodness-of-Fit Tests
2.5K
A goodness-of-fit test is conducted to determine whether the observed frequency values are statistically similar to the frequencies expected for the dataset. Suppose the expected frequencies for a dataset are equal such as when predicting the frequency of any number appearing when casting a die. In that case, the expected frequency is the ratio of the total number of observations (n) to the number of categories (k).
2.5K

