Related Experiment Video
Updated: Aug 16, 2025

15:00
A Tablet-Based Curriculum-Based Measurement Protocol for Kindergarten Writing
Published on: February 7, 2025
703
Potential scoring and predictive bias in interim and summative writing assessments
Deborah K Reed1, Sterett H Mercer2
1Tennessee Reading Research Center, University of Tennessee.
School Psychology (Washington, D.C.)
|December 22, 2022
Summary
Teacher ratings on student writing assessments may introduce bias, particularly for English learners and students needing special education. Masking student identities in evaluations can help reduce scoring disparities and improve fairness in educational assessments.
Area of Science:
- Educational Psychology
- Assessment and Evaluation
- Writing Studies
Background:
- Interim and summative writing assessments inform instructional decisions.
- Potential for rater and score type bias in student writing evaluations is understudied.
- Understanding bias is crucial for equitable educational practices.
Purpose of the Study:
- To investigate rater bias in interim writing assessments.
- To compare teacher ratings versus researcher ratings (blinded).
- To examine score predictive validity across diverse student groups.
Main Methods:
- Analysis of interim writing assessments and state summative data (N=2,621, Grades 3-11).
- Evaluations by familiar teachers and blinded researchers using analytic rubrics.
- Comparison of scores and predictive validity across demographic groups (gender, English learners, free/reduced-price lunch, special education).
Main Results:
- Teachers scored students higher than blinded researchers.
- Female students scored higher than males; English learners, students eligible for free/reduced-price lunch, and students eligible for special education scored lower.
- Score disparities were smaller with researcher ratings.
- Interim scores predicted state outcomes similarly across groups, but some groups had lower overall state scores.
Conclusions:
- Teacher ratings may introduce bias, disproportionately affecting marginalized student groups.
- Masking student identities in writing assessments can mitigate scoring bias.
- Written composition sections of high-stakes tests may be less biased than multiple-choice sections for historically marginalized students.
More Related Videos
Related Concept Videos
Reliability and Validity
12.8K
Reliability and validity are two important considerations that must be made with any type of data collection. Reliability refers to the ability to consistently produce a given result. In the context of psychological research, this would mean that any instruments or tools used to collect data do so in consistent, reproducible ways.
12.8K
Hindsight Biases
3.5K
Hindsight bias leads you to believe that the event you just experienced was predictable, even though it really wasn’t. In other words, you knew all along that things would turn out the way they did. Can you relate this to the phrase "Hindsight is 20/20" now?
3.5K
Review and Preview
7.7K
In statistics, several tools are used to interpret the data. Measures of central tendency represent the characteristics of the data, such as mean, median, and mode. Additionally, measures of variance like standard deviation and range are used to find the spread of data from the mean. Relative standing measures the distance between data locations. Commonly used measures of relative standings are percentile, z score, and quartiles.
Percentiles are a type of fractile that partition data into...
Percentiles are a type of fractile that partition data into...
7.7K
Introduction to z Scores
445
A z score (or standardized value) is measured in units of the standard deviation. It indicates how many standard deviations the value x is above (to the right of) or below (to the left of) the mean, μ. Values of x that are larger than the mean have positive z scores, and values of x that are smaller than the mean have negative z scores. If x equals the mean, then x has a zero z score. It is important to note that the mean of the z scores is zero, and the standard deviation is one.
z scores...
z scores...
445
Bias
4.8K
Bias refers to any tendency that prevents a question from being considered unprejudiced. In research, bias occurs when one outcome or answer is selected or encouraged over others in sampling or testing. Bias can occur during any research phase, including study design, data collection, analysis, and publication.
In statistics, a sampling bias is created when a sample is collected from a population, and some members of the population are not as likely to be chosen as others (remember, each member...
In statistics, a sampling bias is created when a sample is collected from a population, and some members of the population are not as likely to be chosen as others (remember, each member...
4.8K
Surveys
14.9K
Often, psychologists develop surveys as a means of gathering data. Surveys are lists of questions to be answered by research participants, and can be delivered as paper-and-pencil questionnaires, administered electronically, or conducted verbally. Generally, the survey itself can be completed in a short time, and the ease of administering a survey makes it easy to collect data from a large number of people.
14.9K

