Evaluating the Performance of a ChatGPT Model in Rheumatology Exams

Fadi Hassan1, Basem Hijazi2, Mohammad E Naffaa1

  • 1Department of Rheumatology, Galilee Medical Center, Nahariya, Israel, Azrieli Faculty of Medicine, Bar-Ilan University, Safed, Israel.

Summary

Large language models show promise for rheumatology exams, with a knowledge-augmented model achieving 81% accuracy. However, performance significantly drops on image-based questions, highlighting current limitations.

Related Concept Videos