Related Experiment Video
Updated: Mar 22, 2026

Using Human Differentially Expressed Gene Lists to Perform Downstream Pathway Enrichment Analysis and Target Prioritization
Published on: October 3, 2025
ChatGPT versus UpToDate in Preclinical Medical Education: Cross-Sectional Analysis Using Term Frequency-Inverse
Shankar S Thiru1, Nicholas E Aksu2, Matthew Chiang3
1Georgetown University School of Medicine, 3800 Reservoir Rd, Washington, DC, 20007, United States.
Generative artificial intelligence (AI) tools like ChatGPT show moderate alignment with UpToDate for preclinical medical questions, particularly in pharmacology. While not a replacement for evidence-based resources, AI can be a supplementary learning tool for students.
Area of Science:
- Medical Education
- Artificial Intelligence in Medicine
- Preclinical Medical Training
Background:
- Generative AI tools, including ChatGPT, are increasingly adopted by medical students for self-directed learning.
- The reliability of these AI models as supplementary resources for preclinical education is not well-established.
- There is a lack of comparative studies between AI-generated content and established evidence-based medical references like UpToDate.
Purpose of the Study:
- To evaluate the similarity between responses generated by ChatGPT (GPT-4o mini) and UpToDate for preclinical medical education questions.
- To assess the potential of ChatGPT as an adjunctive learning tool in preclinical medical education.
Main Methods:
- A cross-sectional comparison study was conducted using 150 preclinical medical questions.
- ChatGPT responses were generated across 10 separate sessions per question.
- Responses were preprocessed and compared with UpToDate content using TF-IDF cosine similarity, with comparisons against randomized text to assess nonrandom overlap.
Main Results:
- ChatGPT responses showed statistically significant similarity to UpToDate in 59.3% of questions.
- Pharmacology exhibited the highest concordance (mean cosine similarity 0.338), followed by pathology, biochemistry, microbiology, and immunology.
- All subject-level similarity scores surpassed those from randomized text, indicating nonrandom overlap.
Conclusions:
- ChatGPT (GPT-4o mini) demonstrates moderate, meaningful alignment with UpToDate for preclinical topics, excelling in fact-based subjects like pharmacology.
- While not a substitute for evidence-based resources, ChatGPT can function as an accessible supplementary learning tool for medical students.
- Responsible integration of AI in preclinical learning necessitates AI literacy training for critical appraisal and effective use.
More Related Videos
03:37Author Spotlight: Impact of Intergenic Interactions on Disease-Identifying Dark Biomarkers
Published on: March 1, 2024
05:47Evidence-based Knowledge Synthesis and Hypothesis Validation: Navigating Biomedical Knowledge Bases via Explainable AI and Agentic Systems
Published on: June 13, 2025
Related Concept Videos
Improving Translational Accuracy
Improving Translational Accuracy
Multiple Comparison Tests
It would be easy to compare two samples using a significance alpha level of 0.05. In other words, there is only one sample pair to be compared. However, it would be difficult to identify a significantly different sample if the number...
Preclinical Development: Overview
Bioequivalence: Overview
Clinical Trials: Overview