Related Experiment Video
Updated: Jun 23, 2026

10:39
Qualitative and Quantitative Validation of Tools with Rating Scales Aimed at Assessing the Quality of University Service-Learning
Published on: August 29, 2025
Evaluation of five guidelines for option development in multiple-choice item-writing
Rafael J Martínez1, Rafael Moreno, Irene Martín
1Universidad de Sevilla, Facultad de Psicología, Sevilla, Spain.
Psicothema
|May 1, 2009
Summary
Heterogeneous content in multiple-choice test items harms item discrimination. However, "All of the above" options can improve discrimination when used correctly in educational assessments.
Area of Science:
- Educational Measurement
- Psychometrics
Background:
- Effective multiple-choice test item writing is crucial for accurate assessment of student knowledge.
- Previous guidelines suggested avoiding certain item characteristics, such as varying option lengths or content heterogeneity.
Purpose of the Study:
- To empirically evaluate common guidelines for writing multiple-choice test items.
- To determine the impact of specific item characteristics on test performance and discrimination.
Main Methods:
- Analysis of response data from 5013 subjects across 630 multiple-choice items from 21 university achievement tests.
- Statistical examination of item discrimination based on variations in option content, length, and the use of 'None of the above' or 'All of the above' options.
Main Results:
- Heterogeneous content among options negatively impacts item discrimination.
- The 'None of the above' option, when correct, also shows a detrimental effect on discrimination.
- No significant negative effects were found for differing option lengths or the use of specific determiners.
- The 'All of the above' option, when correct, decreased item difficulty and improved discrimination.
Conclusions:
- Educators should avoid heterogeneous content in multiple-choice options to enhance item quality.
- The 'All of the above' option can be a beneficial strategy for improving item discrimination in achievement tests.
Related Concept Videos
Guidelines for Writing Outcome
When developing expected outcomes for a patient care plan, the nurse should adhere to the following recommendations:
Patient outcomes reflect the patient's response to the goal rather than what the nurse aims to achieve. Terminology should be observable and measurable to avoid the reader's interpretation. The desired outcome should be realistic and achievable in the designated care timeframe. Expected outcomes should align with adjunctive therapies. The outcome should enhance care evaluation by...
Patient outcomes reflect the patient's response to the goal rather than what the nurse aims to achieve. Terminology should be observable and measurable to avoid the reader's interpretation. The desired outcome should be realistic and achievable in the designated care timeframe. Expected outcomes should align with adjunctive therapies. The outcome should enhance care evaluation by...
Decision Making: P-value Method
The process of hypothesis testing based on the P-value method includes calculating the P- value using the sample data and interpreting it.
First, a specific claim about the population parameter is proposed. The claim is based on the research question and is stated in a simple form. Further, an opposing statement to the claim is also stated. These statements can act as null and alternative hypotheses: a null hypothesis would be a neutral statement while the alternative hypothesis can have a...
First, a specific claim about the population parameter is proposed. The claim is based on the research question and is stated in a simple form. Further, an opposing statement to the claim is also stated. These statements can act as null and alternative hypotheses: a null hypothesis would be a neutral statement while the alternative hypothesis can have a...
Group Design
The most basic experimental design involves two groups: the experimental group and the control group. The two groups are designed to be the same except for one difference— experimental manipulation. The experimental group gets the experimental manipulation—that is, the treatment or variable being tested—and the control group does not. Since experimental manipulation is the only difference between the experimental and control groups, we can be sure that any differences between the two are due to...
Methods of Medium Optimization
Optimizing growth media enhances microbial proliferation and maximizes product yield. Statistical experimental design methodologies provide structured and reproducible approaches, offering progressively higher levels of robustness and efficiency.The One-Factor-at-a-Time (OFAT) MethodThe One-Factor-at-a-Time (OFAT) method involves adjusting a single variable while keeping all others constant. However, it cannot detect interactions between variables, often leading to suboptimal outcomes when...
Decision Making: Traditional Method
The process of hypothesis testing based on the traditional method includes calculating the critical value, testing the value of the test statistic using the sample data, and interpreting these values.
First, a specific claim about the population parameter is decided based on the research question and is stated in a simple form. Further, an opposing statement to this claim is also stated. These statements can act as null and alternative hypotheses, out of which a null hypothesis would be a...
First, a specific claim about the population parameter is decided based on the research question and is stated in a simple form. Further, an opposing statement to this claim is also stated. These statements can act as null and alternative hypotheses, out of which a null hypothesis would be a...
Multiple Comparison Tests
Multiple comparison test, abbreviated as MCT, is a post hoc analysis generally performed after comparing multiple samples with one or more tests. An MCT will help identify a significantly different sample among multiple samples or a factor among multiple factors.
It would be easy to compare two samples using a significance alpha level of 0.05. In other words, there is only one sample pair to be compared. However, it would be difficult to identify a significantly different sample if the number...
It would be easy to compare two samples using a significance alpha level of 0.05. In other words, there is only one sample pair to be compared. However, it would be difficult to identify a significantly different sample if the number...

