Artificial Intelligence in Medical Assessment: Reliability and Performance of Multimodal Large Language Models in a

Ibrahim Güler1,2,3, Gerrit Grieb2,4, Armin Kraus1

  • 1Department of Plastic, Aesthetic and Hand Surgery, Otto-von-Guericke University, 39120 Magdeburg, Germany.

Summary

Large language models (LLMs) show high accuracy and reproducibility in medical licensing exams. These AI tools demonstrate reliable performance, supporting their use in assessment research.

Related Concept Videos