Related Experiment Video
Updated: May 5, 2026

Author Spotlight: Segmentation and VR for Advanced Neurovascular Interventions
Published on: April 5, 2024
ChatGPT versus human authors: A comparative study of concept maps for clinical reasoning training with virtual
Renata Szydlak1, Yavuz Selim Kiyak2, Inga Hege3
1Department of Bioinformatics and Telemedicine, Jagiellonian University Medical College, Kraków, Poland.
Purpose:
This study investigates whether ChatGPT can generate clinically accurate and pedagogically valuable maps for clinical reasoning (CR) training. The aim is to assess its potential as a tool for supporting the creation of high-quality educational resources for CR training.
Materials And Methods:
We selected 10 diverse virtual patients (VPs) from the European iCoViP project. For each case, CR concept maps were generated by a custom ChatGPT model and compared to expert-created maps available in the CASUS VP system. The comparison encompassed structural metrics (number of concepts, connections, and graph density), clinical content quality (clinical expert evaluation of concept and connection validity), and pedagogical utility (medical educator assessment of clarity, abstraction, and progression). Statistical analysis included Student's t-tests and interrater reliability using weighted Cohen's kappa.
Results:
ChatGPT-generated maps contained significantly more concepts and connections than expert maps, indicating higher structural complexity (p < 0.001), though graph density did not differ significantly. Clinician evaluations showed comparable clinical content quality across both groups, with no statistically significant differences in concept or connection ratings. The educational review revealed that while ChatGPT maps offered comprehensive information, they lacked abstraction, prioritization, and contextual alignment, occasionally exceeding the optimal cognitive load for learners.
Conclusions:
ChatGPT can reliably generate concept maps that match expert-level clinical accuracy. However, limitations in educational clarity and usability underscore the need for expert refinement. With appropriate oversight, large language models (LLMs) such as ChatGPT can support efficient development of learning resources for CR education.
More Related Videos
05:04Author Spotlight: Evaluating Clinicians' Adoption of Ultrasound-Guided Vascular Cannulation Through Simulation Training
Published on: August 9, 2024
05:47Evidence-based Knowledge Synthesis and Hypothesis Validation: Navigating Biomedical Knowledge Bases via Explainable AI and Agentic Systems
Published on: June 13, 2025
Related Concept Videos
Inductive Reasoning
Inductive reasoning is common in descriptive science. A life scientist makes observations and records them. This data can be qualitative or...
Deductive Reasoning
For example, a researcher can deduce specific predictions...
Patient-centered Care