Comparative Assessment of Otolaryngology Knowledge Among Large Language Models

Dante J Merlino1, Santiago R Brufau1, George Saieed1

  • 1Department of Otolaryngology-Head and Neck Surgery, Mayo Clinic, Rochester, Minnesota, U.S.A.

The Laryngoscope
|September 21, 2024
PubMed
Summary

OpenAI's GPT-4 demonstrated superior performance in answering otolaryngology questions compared to other large language models. Prompting for reasoning improved accuracy across all evaluated models, highlighting potential for AI in medical education.