Evaluating large language models vs residents in cataract and refractive surgery: comparative analysis using the

Avi Wallerstein1, Taanvee Ramnawaz, Mathieu Gauvin

  • 1From the Department of Ophthalmology and Visual Sciences, McGill University, Montreal, Quebec, Canada (Wallerstein, Gauvin); LASIK MD, Montreal, Quebec, Canada (Wallerstein, Ramnawaz, Gauvin); School of Medicine, University of Montreal, Montreal, Quebec, Canada (Ramnawaz).

Summary

ChatGPT-4o demonstrated superior accuracy in answering cataract and refractive surgery questions, significantly outperforming ophthalmology residents and other large language models (LLMs). Prompt complexity did not impact LLM performance.