Agreement Between Reasoning-Oriented Generative AI Models and Clinical Educators in Evaluating Japanese Objective

Takanobu Hirosawa1, Masashi Yokose1, Tetsu Sakamoto1

  • 1Department of Diagnostic and Generalist Medicine, Dokkyo Medical University, 880 Kitakobayashi, Mibu-cho, Shimotsuga, Tochigi, 321-0293, Japan, 81 282-87-2498.

Summary

Generative artificial intelligence (GenAI) models showed poor agreement and lower scores compared to clinical educators for evaluating Japanese medical interviews. These AI tools are not yet suitable as standalone evaluators for Objective Structured Clinical Examination transcripts.

Related Concept Videos