Related Experiment Video
Updated: May 11, 2026

14:45
Technical Detail for Robot Assisted Pancreaticoduodenectomy
Published on: September 28, 2019
15.6K
The Evolving Role of ChatGPT (Chat-Generative Pre-Trained Transformer) in General Surgery: A Systematic Review
Carlos Andre Balthazar da Silveira1, Ana Caroline Dias Rasador1, Raquel Nogueira2
1Department of Surgery, Dignity Health St. Joseph's Hospital, Phoenix, Arizona, USA.
Journal of Laparoendoscopic & Advanced Surgical Techniques. Part A
|February 18, 2026
Summary
Chat Generative Pre-Trained Transformer (ChatGPT), a large language model, shows promise in general surgery education, clinical practice, and research. Human oversight remains essential for validating its outputs and ensuring patient safety.
Area of Science:
- Medical Artificial Intelligence
- Natural Language Processing in Surgery
- Large Language Models (LLMs)
Background:
- Large language models (LLMs) like Chat Generative Pre-Trained Transformer (ChatGPT) are increasingly accessible with potential medical applications.
- While LLM utility is explored in various surgical fields, its specific impact on general surgery requires systematic evaluation.
- This review assesses the current evidence for ChatGPT's educational, clinical, and research applications in general surgery.
Purpose of the Study:
- To systematically review and evaluate the existing literature on the applications of ChatGPT in general surgery.
- To analyze ChatGPT's performance across educational, clinical, and research domains within general surgery.
- To identify the strengths and limitations of ChatGPT in the context of general surgery practice and education.
Main Methods:
- A comprehensive systematic search was conducted across major databases (PubMed, Cochrane Central, Scopus, SciELO, LILACS) up to December 2023.
- Studies evaluating ChatGPT's utility in general surgery's educational, research, and clinical domains were included.
- Both analytic and descriptive studies were considered, excluding studies on other AI platforms and conference abstracts.
Main Results:
- Twenty-three studies demonstrated ChatGPT's broad applicability in general surgery, covering disease Q&A, clinical practice, education, and research.
- ChatGPT showed high accuracy (up to 87%) for colorectal surgical questions and favorable patient-facing responses, particularly in bariatric and transplant surgery.
- Performance varied in clinical decision-making (0-86.7%) and board exam assessments (48-76.4%), with limitations in referencing; human validation is crucial.
Conclusions:
- ChatGPT exhibits significant potential across general surgery education, clinical practice, and research.
- Outputs from ChatGPT require continuous human oversight and expert validation for safe and effective implementation.
- Further research is needed to refine ChatGPT's accuracy and reliability in specific surgical contexts.
More Related Videos
Related Concept Videos
Improving Translational Accuracy
Base complementarity between the three base pairs of mRNA codon and the tRNA anticodon is not a failsafe mechanism. Inaccuracies can range from a single mismatch to no correct base pairing at all. The free energy difference between the correct and nearly correct base pairs can be as small as 3 kcal/ mol. With complementarity being the only proofreading step, the estimated error frequency would be one wrong amino acid in every 100 amino acids incorporated. However, error frequencies observed in...
Improving Translational Accuracy
Base complementarity between the three base pairs of mRNA codon and the tRNA anticodon is not a failsafe mechanism. Inaccuracies can range from a single mismatch to no correct base pairing at all. The free energy difference between the correct and nearly correct base pairs can be as small as 3 kcal/ mol. With complementarity being the only proofreading step, the estimated error frequency would be one wrong amino acid in every 100 amino acids incorporated. However, error frequencies observed in...

