Comparing Scoring Consistency of Large Language Models with Faculty for Formative Assessments in Medical Education

Radhika Sreedhar1, Linda Chang2, Ananya Gangopadhyaya2

  • 1University of Illinois College of Medicine, Chicago, IL, USA. sreedhar@uic.edu.

PubMed
Summary

Large language models (LLMs) show promise in assessing medical student critical appraisal assignments, offering consistent feedback and reducing faculty time. This study found LLMs comparable to faculty grading, highlighting their potential in medical education.