Related Experiment Videos

Know when to trust: Making AI scoring more reliable for educational assessment

Peter Organisciak1, Selcuk Acar2

  • 1University of Denver, 1999 E Evans Ave, Denver, CO, 80208, USA. peter.organisciak@du.edu.

Summary

This study enhances automated scoring using large language models (LLMs) with three new methods. These improvements increase the trustworthiness and accuracy of LLM-based educational scoring tools.