Estimating Between-Person and Within-Person Subscore Reliability with Profile Analysis.
Okan Bulut1, Mark L Davison2, Michael C Rodriguez2
1a Centre for Research in Applied Measurement and Evaluation , University of Alberta.
Multivariate Behavioral Research
|November 30, 2016
Summary
This study introduces a new profile reliability approach for educational and psychological testing subscores. It reveals a trade-off between within-person and between-person reliability, crucial for understanding distinct subscore value.
Area of Science:
- Educational Measurement
- Psychometrics
- Psychological Testing
Background:
- Subscores are increasingly important for diagnosing examinee strengths and weaknesses.
- Previous research primarily focused on individual subscore reliability, neglecting their distinctiveness and added value over total scores.
Purpose of the Study:
- To introduce a novel profile reliability approach for subscores.
- To partition overall subscore reliability into within-person and between-person components.
- To evaluate the distinctiveness and added value of subscores.
Main Methods:
- Developed and demonstrated a profile reliability approach.
- Estimated within-person and between-person reliability coefficients.
- Utilized simulation and real data studies with various scoring methods (e.g., IRT).
Main Results:
- A significant trade-off exists between within-person and between-person subscore reliability.
- The proposed profile reliability coefficients effectively quantify distinctness under different testing conditions.
- Subtest length, correlations, and number of subtests impact reliability.
Conclusions:
- The profile reliability approach offers a more comprehensive evaluation of subscores.
- Understanding the balance between within-person and between-person reliability is key for test development.
- This method aids in determining the diagnostic utility of subscores in various contexts.
Related Concept Videos
Reliability and Validity
14.2K
Reliability and validity are two important considerations that must be made with any type of data collection. Reliability refers to the ability to consistently produce a given result. In the context of psychological research, this would mean that any instruments or tools used to collect data do so in consistent, reproducible ways.
14.2K
Variation
8.2K
An important characteristic of any set of data is the variation in the data. In some data sets, the data values are concentrated closely near the mean; in other data sets, the data values are more widely spread out from the mean. The most common measure of variation, or spread, is the standard deviation, which is the square root of variance.
When independent and dependent variables are plotted on a scatter plot, the slope of a line is a value that describes the rate of change between the two...
When independent and dependent variables are plotted on a scatter plot, the slope of a line is a value that describes the rate of change between the two...
8.2K
Confidence Coefficient
10.8K
The confidence coefficient is also known as the confidence level or degree of confidence. It is the percent expression for the probability, 1-α, that the confidence interval contains the true population parameter assuming that the confidence interval is obtained after sufficient unbiased sampling; for example, if the CL = 90%, then in 90 out of 100 samples the interval estimate will enclose the true population parameter. Here α is the area under the curve, distributed equally under...
10.8K
One-Way ANOVA: Equal Sample Sizes
4.3K
One-Way ANOVA can be performed on three or more samples with equal or unequal sample sizes. When one-way ANOVA is performed on two datasets with samples of equal sizes, it can be easily observed that the computed F statistic is highly sensitive to the sample mean.
Different sample means can result in different values for the variance estimate: variance between samples. This is because the variance between samples is calculated as the product of the sample size and the variance between the...
Different sample means can result in different values for the variance estimate: variance between samples. This is because the variance between samples is calculated as the product of the sample size and the variance between the...
4.3K
Spearman's Rank Correlation Test
1.6K
Spearman's rank correlation test, also known as Spearman's rho, is a nonparametric method for assessing the strength and direction of association between two variables. This test is particularly valuable when the data distribution is unknown or when the assumption of normality does not hold. Named after the English psychologist and statistician Dr. Charles Edward Spearman, it serves as the nonparametric counterpart to Pearson's correlation coefficient.
Spearman's test calculates correlation by...
Spearman's test calculates correlation by...
1.6K
Self-Report Tests of Personality
1.1K
Self-report inventories are objective personality assessments that use multiple-choice items or numbered scales, typically ranging from 1 (strongly disagree) to 5 (strongly agree). They are often called Likert scales after Rensis Likert. These inventories are widely used due to their ease of administration and cost-effectiveness. One of the most prominent examples is the Minnesota Multiphasic Personality Inventory (MMPI), initially developed in the 1940s to assess abnormal personality traits.
1.1K


