Accuracy of Large Language Models in Answering Dental Examination Questions: A Systematic Review and Meta-Analysis

Mahmood Dashti1, Farshad Khosraviani2, Atieh Meyari3

  • 1Dentofacial Deformities Research Center, Research Institute of Dental Sciences, Shahid Beheshti University of Medical Sciences, Tehran, Iran; Department of Artificial Intelligence Engineering, Graduate School of Natural and Applied Sciences, Istinye University, Istanbul, Türkiye.

Summary

Large language models (LLMs) show moderate accuracy on dental exam questions, with ChatGPT-4 and Copilot performing best. These AI tools are not yet suitable for independent clinical decisions but can aid dental education.