Related Experiment Video
Updated: May 1, 2026

Augmenting Large Language Models via Vector Embeddings to Improve Domain-Specific Responsiveness
Published on: December 6, 2024
Evaluating data extraction error by a large language model from randomised controlled trials: a large-scale empirical
Shiqi Fan1, Ming Chen2, Suhail A Doi3
1Proof of Concept Center, Shanghai Eastern Hepatobiliary Surgery Hospital, Shanghai, China.
Large language models (LLMs) like Claude 3.5 Sonnet show low data extraction error rates from randomized controlled trials (RCTs). However, careful verification of LLM outputs is crucial for evidence synthesis applications.
Area of Science:
- Medical informatics
- Artificial intelligence in healthcare
- Clinical trial methodology
Background:
- Large language models (LLMs) offer potential for automating data extraction from clinical research.
- Evaluating the accuracy of LLMs in extracting data from randomized controlled trials (RCTs) is essential.
Purpose of the Study:
- To assess the data extraction accuracy of Claude 3.5 Sonnet on RCTs.
- To identify common error types and influencing factors in LLM-based data extraction.
Main Methods:
- An empirical study compared Claude 3.5 Sonnet's extractions against a human-verified dataset of 664 RCTs.
- Data extraction focused on basic trial information and adverse outcomes.
- Error rates were calculated and analyzed by error type and trial reporting quality (CONSORT adherence).
Main Results:
- The overall data extraction error rate for Claude 3.5 Sonnet was 6.6%.
- Misallocation (57.1%) and omitted data (23.2%) were the most frequent error types.
- Higher adherence to Consolidated Standards of Reporting Trials (CONSORT) guidelines correlated with lower LLM extraction errors.
Conclusions:
- Claude 3.5 Sonnet demonstrates a relatively low error rate for RCT data extraction.
- LLM applications in evidence synthesis require rigorous human oversight and detailed checking of outputs.
More Related Videos
05:47Evidence-based Knowledge Synthesis and Hypothesis Validation: Navigating Biomedical Knowledge Bases via Explainable AI and Agentic Systems
Published on: June 13, 2025
08:05Measuring Statistical Learning Across Modalities and Domains in School-Aged Children Via an Online Platform and Neuroimaging Techniques
Published on: June 30, 2020
Related Concept Videos
Improving Translational Accuracy
Improving Translational Accuracy
Systematic Error: Methodological and Sampling Errors
Sampling errors originate from improper sampling methods or the wrong sample population. These errors can be minimized by refining the sampling strategy. Defective instruments or faulty calibrations are the sources of instrumental...