Related Experiment Video
Updated: May 4, 2026

Noninvasive, In-pen Approach Test for Laboratory-housed Pigs
Published on: June 5, 2019
Chat-GPT in triage: Still far from surpassing human expertise - An observational study
Arian Zaboli1, Francesco Brigo1, Gloria Brigiari2
1Innovation, Research and Teaching Service (SABES-ASDAA), Teaching Hospital of the Paracelsus Medical Private University (PMU), Bolzano, Italy.
Background:
Triage is essential in emergency departments (EDs) to prioritize patient care based on clinical urgency. Recent investigations have explored the role of large language models (LLMs) in triage, but their effectiveness compared to human triage remains uncertain. This study assessed the effectiveness of ChatGPT 4.0 in triaging ED patients.
Methods:
This retrospective study analyzed data from 2658 patients. Triage codes assigned by human triage personnel were compared with those assigned by Artificial Intelligence (AI) triage using Chat-GPT 4.0. Agreement between human and AI triage was assessed using Cohen's kappa statistic. Clinical outcomes were evaluated through Receiver Operating Characteristic (ROC) curves to determine predictive accuracy. Sensitivity and specificity of both triage systems were compared across different symptoms using 2 × 2 contingency tables.
Results:
The Cohen's kappa statistic for agreement between human and AI triage was 0.125 (95 % CI: 0.100-0.134). ROC analysis demonstrated that human triage outperformed AI in predicting all study outcomes, with statistically significant differences. For 30-day mortality, the ROC of human triage was 0.88, while for AI triage it was 0.70, p < 0.001. A similar result was observed for life-saving interventions, where human triage had an ROC of 0.98 and AI triage 0.87, p = 0.014. For specific symptoms, human triage showed superior sensitivity and specificity.
Conclusions:
LLMs like Chat-GPT 4.0 have limited utility in ED triage, particularly due to their lower sensitivity for high-risk patients, which lead to under-triage. Human triage remains more reliable than Chat-GPT.
More Related Videos
06:28E-Patient Counseling Trial E-PACO: Computer Based Education versus Nurse Counseling for Patients to Prepare for Colonoscopy
Published on: August 1, 2019
05:39Author Spotlight: A Non-Intubated Video-Assisted Thoracoscopic Surgery with Multimodal Analgesia and Sevoflurane Inhalation Anesthesia
Published on: May 26, 2023
Related Concept Videos
Case Studies
Naturalistic Observations
Group Polarization
Higher Mental Functions of the Brain: Language
Language formation and comprehension take place in the dominant hemisphere. The dominant hemisphere is responsible for understanding the meaning of spoken, written, or sign language, as well as the ability to communicate. For most people, the left hemisphere is the dominant one. The right hemisphere, then, gives tone and emotional context to the...
Extrasensory Perception
Precognition involves foreseeing future events, such as predicting an accident before it happens. An example of precognition could be someone dreaming about a specific event, like a car crash, which then occurs...