Related Experiment Video
Updated: May 7, 2025

Objectification of Tongue Diagnosis in Traditional Medicine, Data Analysis, and Study Application
Published on: April 14, 2023
Evaluation of the ability of large language models to self-diagnose oral diseases
Shiyang Zhuang1,2,3, Yuanhao Zeng4, Shaojunjie Lin3
1Department of Stomatology, the First Affiliated Hospital, Fujian Medical University, Fuzhou 350005, China.
Abstract:
Large language models (LLMs) offer potential in primary dental care. We conducted an evaluation of LLMs' diagnostic capabilities across various oral diseases and contexts. All LLMs showed diagnostic capabilities for temporomandibular joint disorders, periodontal disease, dental caries, and malocclusion. The prompts did not affect the performance of ChatGPT 3.5. When Chinese was used, the diagnostic ability of ChatGPT 3.5 for pulpitis improved (0% vs. 61.7%, p < 0.001), while the ability to diagnose pericoronitis decreased (8% vs. 0%, p < 0.001). For ChatGPT 4.0 in Chinese, they were both improved (0% vs. 92%, 8% vs. 72%, p < 0.001, respectively). Claude 2 exhibited the highest accuracy in diagnosing pulpitis (36%, p = 0.048), ChatGPT 4.0 showed complete diagnostic capability for pericoronitis. Llama 2 and Claude 3.5 Sonnet exhibited complete diagnostic capability for oral cancer. In conclusion, LLMs may be a potential tool for daily dental care but need further updates.
Related Concept Videos
Assessment of the Mouth
Mouth Inspection
The inspection begins with visually examining the mouth for symmetry, color, and size.
Oral Cavity
Teeth: The teeth are the hardest structures in our bodies. Humans have two sets of teeth throughout their lifetime: deciduous (baby) teeth and permanent teeth. Each tooth consists of several parts: the crown (visible part), the root (embedded in the jaw), enamel (hard outer...
Language and Cognition

