Related Experiment Video
Updated: Jan 9, 2026

The Dyspepsia Educational Tool As a Novel Aid in Dyspepsia Management
Published on: June 29, 2019
Readability of Professional Medical Content on Acute Appendicitis: A Comparative Cross-Sectional Study of ChatGPT and
Maira Jalil1, Saow Renn Ding2, Nikita M Talpallikar3
1Internal Medicine, University of Debrecen, Faculty of Medicine, Debrecen, HUN.
Abstract:
Introduction Acute appendicitis is one of the most common diseases occurring due to inflammation of the vermiform appendix, which requires surgical intervention. With the advent of standardized artificial intelligence (AI) tools such as ChatGPT (OpenAI, San Francisco, CA), AI-based search engines have emerged as a secondary means for patients to educate themselves about their health. The readability of each response is important for concept understanding as well as the impact of novel therapeutics. Aims This study aims to evaluate and compare the readability of medical information on acute appendicitis generated by an AI language model and UpToDate (Wolters Kluwer Health, Waltham, MA) using established readability metrics. Methodology A comparative cross-sectional study was conducted to evaluate the readability of six ChatGPT-4o and six UpToDate responses on acute appendicitis. Readability parameters were assessed using WebFX (WebFX®, Harrisburg, PA), and differences between sources were analyzed using the Mann-Whitney U test in IBM SPSS Statistics software, version 25 (IBM Corp., Armonk, NY) and R (v4.3.2, The R Core Team, R Foundation for Statistical Computing, Vienna, Austria). Results UpToDate had a higher word count and higher words per sentence than ChatGPT (both p < 0.05). ChatGPT had a lower absolute difficult-word count (p = 0.002) but a higher difficult-word percentage (p = 0.002). Differences in Flesch Reading Ease (FRE), Flesch-Kincaid Grade Level (FKGL), Simple Measure of Gobbledygook (SMOG), and sentence count were not statistically significant (all p > 0.05). Conclusions ChatGPT produced more concise content compared to UpToDate, but its higher proportion of difficult words may limit comprehension, highlighting the need to balance brevity with readability in AI-generated medical information.
More Related Videos
06:28E-Patient Counseling Trial E-PACO: Computer Based Education versus Nurse Counseling for Patients to Prepare for Colonoscopy
Published on: August 1, 2019
05:50Author Spotlight: Point-of-Care Ultrasound for Gastric Content Assessment and Risk Stratification in Perioperative Care
Published on: September 22, 2023
Related Concept Videos
Appendicitis-II: Diagnostic Studies and Management
Diagnosing Appendicitis
It requires a multifaceted approach, starting with a detailed physical examination to pinpoint the location and nature of the pain and identify any associated symptoms. Laboratory tests play a crucial role. A complete Blood Count (CBC) typically reveals leukocytosis (an increased number of...
Appendicitis-I: Introduction
Etiology: Appendicitis can arise from various causes, primarily rooted in the obstruction of the appendix lumen. Factors contributing to this obstruction include fecal accumulation, lymphoid hyperplasia and, in...
Urinary Tract Infection III: Diagnostic Studies and Interprofessional Care
Acute Pyelonephritis II: Diagnostic Studies and Management
Acute Pancreatitis II: Clinical Manifestations and Management
Imaging Studies V: Intravenous Urography and Retrograde Pyelography