Jove
Visualize
Contact Us
JoVE
x logofacebook logolinkedin logoyoutube logo
ABOUT JoVE
OverviewLeadershipBlogJoVE Help Center
AUTHORS
Publishing ProcessEditorial BoardScope & PoliciesPeer ReviewFAQSubmit
LIBRARIANS
TestimonialsSubscriptionsAccessResourcesLibrary Advisory BoardFAQ
RESEARCH
JoVE JournalMethods CollectionsJoVE Encyclopedia of ExperimentsArchive
EDUCATION
JoVE CoreJoVE BusinessJoVE Science EducationJoVE Lab ManualFaculty Resource CenterFaculty Site
Terms & Conditions of Use
Privacy Policy
Policies

Related Concept Videos

Interval Level of Measurement00:55

Interval Level of Measurement

17.8K
For effective statistical analysis, data are classified into four levels of measurement—nominal, ordinal, interval, and ratio.
Data measured using the interval scale are similar to ordinal level data because they have a definite arrangement. However, in the interval level of measurement, the differences between data values are meaningful even though the data does not have a starting point.
Temperature is measured using the interval scale. It is measurable data, and the difference between...
17.8K
Self-Report Tests of Personality01:22

Self-Report Tests of Personality

675
Self-report inventories are objective personality assessments that use multiple-choice items or numbered scales, typically ranging from 1 (strongly disagree) to 5 (strongly agree). They are often called Likert scales after Rensis Likert. These inventories are widely used due to their ease of administration and cost-effectiveness. One of the most prominent examples is the Minnesota Multiphasic Personality Inventory (MMPI), initially developed in the 1940s to assess abnormal personality traits.
675
Ordinal Level of Measurement00:55

Ordinal Level of Measurement

31.3K
The way a set of data is measured is called its level of measurement. Correct statistical procedures depend on a researcher being familiar with levels of measurement. For analysis, data are classified into four levels of measurement—nominal, ordinal, interval, and ratio.
Data measured using an ordinal scale are similar to nominal scale data, but there is one major difference. The ordinal scale data can be ordered. An example of ordinal scale data is a list of the top five national parks...
31.3K
Ratio Level of Measurement00:54

Ratio Level of Measurement

20.5K
The way a set of data is measured is called its level of measurement. Correct statistical procedures depend on a researcher being familiar with levels of measurement. For analysis, data are classified into four levels of measurement—nominal, ordinal, interval, and ratio.
A set of data measured using the ratio scale takes care of the ratio problem and provides complete information. Ratio scale data are like interval scale data, except they have a zero point and ratios can be calculated....
20.5K
Interpretation of Confidence Intervals01:19

Interpretation of Confidence Intervals

9.0K
A confidence interval is a better estimate of the population than a point estimate, as it uses a range of values from a sample instead of a single value.
Confidence intervals have confidence coefficients that are crucial for their interpretation. The most common confidence coefficients are 0.90, 0.95, and 0.99, which can be written as percentages–90%, 95%, and 99%, respectively.
Suppose a person calculates a confidence interval with a confidence coefficient of 0.95. In that case, they can...
9.0K
Statistical Analysis: Overview01:11

Statistical Analysis: Overview

13.7K
When we take repeated measurements on the same or replicated samples, we will observe inconsistencies in the magnitude. These inconsistencies are called errors. To categorize and characterize these results and their errors, the researcher can use statistical analysis to determine the quality of the measurements and/or suitability of the methods.
One of the most commonly used statistical quantifiers is the mean, which is the ratio between the sum of the numerical values of all results and the...
13.7K

You might also read

Related Articles

Articles linked to this work by shared authors, journal, and citation graph.

Sort by
Same author

How to Efficiently Recruit Participants in a Perinatal Cohort Study; CAN-B Cohort Strategies/Lessons Learned.

Journal of obstetrics and gynaecology Canada : JOGC = Journal d'obstetrique et gynecologie du Canada : JOGC·2026
Same author

Interactive, Personalized Patient Decision Aid for COVID-19 Vaccination in Canada: User-Centered Design Approach.

JMIR human factors·2026
Same author

Global Variation in Anaphylaxis Treatment: A Scoping Review Protocol.

Clinical and experimental allergy : journal of the British Society for Allergy and Clinical Immunology·2026
Same author

Unheard Voices from Earth: A Fable from 2184.

Medical education·2025
Same author

Serum sickness-like reactions to amoxicillin in children: Drug provocation test duration, recurrence, cross-reactivity.

Pediatric allergy and immunology : official publication of the European Society of Pediatric Allergy and Immunology·2025
Same author

CMV primary and non-primary infections among daycare workers, and development of strategies to prevent infection (EDUQ-CMV): a mixed-method study protocol.

BMC infectious diseases·2025

Related Experiment Video

Updated: Dec 22, 2025

Qualitative and Quantitative Validation of Tools with Rating Scales Aimed at Assessing the Quality of University Service-Learning
10:39

Qualitative and Quantitative Validation of Tools with Rating Scales Aimed at Assessing the Quality of University Service-Learning

Published on: August 29, 2025

914

Accuracy of rating scale interval values used in multiple mini-interviews: a mixed methods study.

Philippe Bégin1,2, Robert Gagnon3, Jean-Michel Leduc3

  • 1Faculty of Medicine, Université de Montréal, Montreal, Canada. philippe.begin@umontreal.ca.

Advances in Health Sciences Education : Theory and Practice
|May 8, 2020
PubMed
Summary

This study established accurate scoring for multiple mini-interviews (MMI) by validating rater intent with quantitative data. This ensures fairer candidate rankings by reflecting the true weight of each score in the Canadian integrated French MMI (IFMMI).

Keywords:
AdmissionBiasEvaluation criteriaGradingIntervalInterviewLikert scaleMMIMedical schoolRatingRubrics

More Related Videos

Author Spotlight: A Novel Setup to Conduct Naturalistic Laboratory Experiments with Real Human Actors in Scenarios
07:43

Author Spotlight: A Novel Setup to Conduct Naturalistic Laboratory Experiments with Real Human Actors in Scenarios

Published on: August 4, 2023

2.6K
A Protocol of Manual Tests to Measure Sensation and Pain in Humans
07:28

A Protocol of Manual Tests to Measure Sensation and Pain in Humans

Published on: December 19, 2016

21.5K

Related Experiment Videos

Last Updated: Dec 22, 2025

Qualitative and Quantitative Validation of Tools with Rating Scales Aimed at Assessing the Quality of University Service-Learning
10:39

Qualitative and Quantitative Validation of Tools with Rating Scales Aimed at Assessing the Quality of University Service-Learning

Published on: August 29, 2025

914
Author Spotlight: A Novel Setup to Conduct Naturalistic Laboratory Experiments with Real Human Actors in Scenarios
07:43

Author Spotlight: A Novel Setup to Conduct Naturalistic Laboratory Experiments with Real Human Actors in Scenarios

Published on: August 4, 2023

2.6K
A Protocol of Manual Tests to Measure Sensation and Pain in Humans
07:28

A Protocol of Manual Tests to Measure Sensation and Pain in Humans

Published on: December 19, 2016

21.5K

Area of Science:

  • Medical Education
  • Psychometrics
  • Assessment and Evaluation

Background:

  • Multiple mini-interviews (MMI) use ordinal rating scales, assuming linear intervals between scores for candidate ranking.
  • This assumption is often unvalidated, potentially introducing systemic bias and distorting final cumulative scores.
  • Accurate interval values are crucial for the validity and fairness of MMI assessments.

Purpose of the Study:

  • To establish validated rating scale values reflecting rater intent in the Canadian integrated French MMI (IFMMI).
  • To quantitatively validate these values using an independent method.
  • To explore the impact of revised scale values on final candidate rankings and understand rater rationale.

Main Methods:

  • A 4-round consensus-group exercise with 42 experienced MMI interviewers to determine relative values for a 6-point scale (A-F).
  • Parallel quantitative validation by comparing average scores assigned to candidates when specific scale options were used over three years.
  • Simulation of the impact of new score values on final rankings using data from 43,412 IFMMI stations and 4,345 applicants.

Main Results:

  • Experienced raters assigned values: B=86.7%, C=69.5%, D=51.2%, E=29.3% (vs. A=100%, F=0%).
  • Quantitative validation yielded similar values: B=87.1%, C=70.4%, D=51.2%, E=31.8%.
  • Qualitative analysis revealed raters perceive lower scores (serious offenses) as carrying more weight than higher scores (minor details).

Conclusions:

  • Consulting experienced interviewers is effective for establishing MMI rating scale values that align with rater intent.
  • Revised scale values, reflecting rater intent, improve the accuracy and fairness of the IFMMI assessment process.
  • The study demonstrated potential shifts in candidate rankings (±21 to +5 percentiles) with the new scale, averaging ±1.4 percentiles.