Related Experiment Video
Updated: Jan 17, 2026

Augmenting Large Language Models via Vector Embeddings to Improve Domain-Specific Responsiveness
Published on: December 6, 2024
Automatically Extracting Numerical Results from Randomized Controlled Trials with Large Language Models
Hye Sun Yun1, David Pogrebitskiy1, Iain J Marshall2
1Northeastern University, Boston, MA, USA.
Large language models (LLMs) show promise for automating meta-analyses by extracting data from randomized controlled trials (RCTs). While effective for simple outcomes, LLMs struggle with complex data requiring inference.
Area of Science:
- Medical Informatics
- Natural Language Processing
- Clinical Trials
Background:
- Meta-analyses are crucial for robust treatment effectiveness estimates, synthesizing findings from multiple randomized controlled trials (RCTs).
- Current meta-analysis requires laborious manual data extraction from individual trial reports, limiting efficiency and scalability.
- Automating this data extraction using language technologies could enable on-demand meta-analyses.
Purpose of the Study:
- To evaluate the capability of modern large language models (LLMs) in reliably extracting numerical findings from clinical trial reports for meta-analysis.
- To assess LLM performance in zero-shot conditional extraction of numerical results linked to interventions, comparators, and outcomes.
Main Methods:
- Development and release of a granular evaluation dataset of clinical trial reports with annotated numerical findings.
- Evaluation of seven large language models (LLMs) using a zero-shot approach on the annotated dataset.
- Focus on extracting numerical results for interventions, comparators, and outcomes from trial reports.
Main Results:
- Massive LLMs demonstrate near-capability for fully automatic meta-analysis, particularly for dichotomous outcomes like mortality.
- LLMs exhibit poor performance when outcome measures are complex and require inferential reasoning to tally results.
- Performance limitations persist even for LLMs trained on biomedical texts.
Conclusions:
- Large language models (LLMs) are approaching the goal of fully automatic meta-analysis of randomized controlled trials (RCTs).
- Current LLMs face significant limitations in extracting and synthesizing complex numerical data from trial reports.
- Further advancements in LLMs are needed to overcome challenges in inferential data processing for comprehensive meta-analysis.
More Related Videos
05:47Evidence-based Knowledge Synthesis and Hypothesis Validation: Navigating Biomedical Knowledge Bases via Explainable AI and Agentic Systems
Published on: June 13, 2025
03:37Author Spotlight: Impact of Intergenic Interactions on Disease-Identifying Dark Biomarkers
Published on: March 1, 2024
Related Concept Videos
Randomized Experiments
Simple randomization
Simple...
Improving Translational Accuracy
Improving Translational Accuracy
Regression Toward the Mean
Types of Biopharmaceutical Studies: Controlled and Non-Controlled Approaches
Non-controlled studies, commonly employed for initial exploration, lack a control group, rendering them susceptible to biases and external influences. In contrast,...
Statistical Software for Data Analysis and Clinical Trials