Evaluating the performance of large language models on the ASPS In-Service Examination: A comparative analysis with

Ramin Shekouhi1, Mary M Holohan1, Oygul Mirzalieva2

  • 1Division of Plastic and Reconstructive Surgery, Department of Surgery, Louisiana State University Health Sciences Center, New Orleans, LA, USA.

Summary

Leading large language models (LLMs) show high accuracy on the Plastic Surgery In-Service Training Examination (PSITE), often surpassing resident and practitioner performance. This indicates their potential in surgical education.

Related Concept Videos