Optimal discrimination index and discrimination efficiency for essay questions
1Office of Educational Services, Faculty of Medicine, The Chinese University of Hong Kong, Shatin, Hong Kong, chan50b@gmail.com.
Summary
Discrimination index guidelines for essay questions differ from multiple-choice tests. Optimal values for essay questions, particularly with a 50% passing mark, range from 0.12 to 0.31, not 0.40 or higher.
Area of Science:
- Educational Measurement
- Psychometrics
Background:
- Standard guidelines for discrimination index are often applied to both multiple-choice and essay questions.
- This indiscriminate application may not be appropriate due to the inherent differences in question formats.
Purpose of the Study:
- To independently derive the optimal discrimination index for essay questions.
- To establish appropriate guidelines for the discrimination index of essay questions.
Main Methods:
- Derived the optimal discrimination index for essay questions under normality conditions.
- Analyzed the satisfactory region for the discrimination index with a 50% passing mark.
- Investigated the relationship between the optimal discrimination index and the range of scores.
- Defined discrimination efficiency.
Main Results:
- The satisfactory region for the discrimination index of essay questions (50% passing mark) is 0.12–0.31.
- Optimal discrimination index for essay questions increases proportionally with the score range.
- Discrimination efficiency was defined as the ratio of observed to optimal discrimination index.
Conclusions:
- Recommended guidelines for the discrimination index of essay questions are provided.
- The study highlights the need for distinct criteria for evaluating essay question quality compared to multiple-choice questions.
Related Concept Videos
Fisher's Exact Test
1.4K
Fisher's exact test is a statistical significance test widely used to analyze 2x2 contingency tables, particularly in situations where sample sizes are small. Unlike the chi-squared test, which approximates P-values and assumes minimum expected frequencies of at least five in each cell, Fisher's exact test calculates the exact probability (P-value) of observing the data or more extreme results under the null hypothesis. This feature makes it especially valuable when the assumptions of...
1.4K
Stereotypes, Prejudice, and Discrimination
77.8K
Humans are very diverse and although we share many similarities, we also have many differences. The social groups we belong to help form our identities (Tajfel, 1974). These differences may be difficult for some people to reconcile, which may lead to prejudice toward people who are different. Prejudice is a negative attitude and feeling toward an individual based solely on one’s membership in a particular social group (Allport, 1954; Brown, 2010). Prejudice is common against people who...
77.8K
Receiver Operating Characteristic Plot
584
A ROC (Receiver Operating Characteristic) plot is a graphical tool used to assess the performance of a binary classification model by illustrating the trade-off between sensitivity (true positive rate) and specificity (false positive rate). By plotting sensitivity against 1 - specificity across various threshold settings, the ROC curve shows how well the model distinguishes between classes, with a curve closer to the top-left corner indicating a more accurate model. The area under the ROC curve...
584
Expected Frequencies in Goodness-of-Fit Tests
7.0K
A goodness-of-fit test is conducted to determine whether the observed frequency values are statistically similar to the frequencies expected for the dataset. Suppose the expected frequencies for a dataset are equal such as when predicting the frequency of any number appearing when casting a die. In that case, the expected frequency is the ratio of the total number of observations (n) to the number of categories (k).
7.0K
Goodness-of-Fit Test
7.1K
The goodness-of-fit test is a type of hypothesis test which determines whether the data "fits" a particular distribution. For example, one may suspect that some anonymous data may fit a binomial distribution. A chi-square test (meaning the distribution for the hypothesis test is chi-square) can be used to determine if there is a fit. The null and alternative hypotheses may be written in sentences or stated as equations or inequalities. The test statistic for a goodness-of-fit test is given as...
7.1K
Statistical Methods to Analyze Parametric Data: Student t-Test and Goodness-of-Fit Test
6.2K
In parametric statistics, two fundamental tests stand out for their utility and wide application: the Student's t-test and goodness-of-fit tests. These tests provide researchers with a robust method for drawing insights from data, testing hypotheses, and making informed decisions based on their findings.
The Student's t-test is a statistical test that examines if there is a statistically significant difference between the means of two groups. This test is instrumental when dealing with...
The Student's t-test is a statistical test that examines if there is a statistically significant difference between the means of two groups. This test is instrumental when dealing with...
6.2K


