Related Experiment Video
Updated: Jan 13, 2026

Augmenting Large Language Models via Vector Embeddings to Improve Domain-Specific Responsiveness
Published on: December 6, 2024
Evaluating the Performance of AI Large Language Models in Detecting Pediatric Medication Errors Across Languages: A
Rana K Abu-Farha1, Haneen Abuzaid2, Jena Alalawneh3
1Clinical Pharmacy and Therapeutics Department, Faculty of Pharmacy, Applied Science Private University, Amman 11937, Jordan.
Microsoft Copilot showed the highest accuracy in detecting pediatric medication errors among four AI models. Performance varied by language, with Arabic generally showing lower accuracy, highlighting the need for better multilingual AI training.
Area of Science:
- Artificial Intelligence in Healthcare
- Pharmacovigilance
- Pediatric Drug Safety
Background:
- Medication errors pose a significant risk in pediatric pharmacotherapy.
- Evaluating the efficacy of artificial intelligence (AI) tools for medication error detection is crucial.
Purpose of the Study:
- To assess the performance of four AI models (GPT-5, GPT-4, Microsoft Copilot, Google Gemini) in identifying medication errors in pediatric case scenarios.
- To compare AI model performance across English and Arabic languages.
Main Methods:
- Sixty pediatric cases, half containing medication errors across four therapeutic systems, were analyzed.
- AI models were tested using a unified prompt in both English and Arabic.
- Performance metrics included accuracy, sensitivity, specificity, and reproducibility, analyzed using SPSS version 22.
Main Results:
- Microsoft Copilot achieved the highest accuracy (86.7% English, 85.0% Arabic), followed by GPT-5.
- Google Gemini exhibited the lowest accuracy (76.7% English, 73.3% Arabic).
- Arabic language performance was generally lower than English; Microsoft Copilot demonstrated superior reproducibility and inter-language agreement.
Conclusions:
- Microsoft Copilot outperformed other AI models in detecting pediatric medication errors in this study.
- The findings underscore the need for enhanced multilingual AI training to ensure equitable performance across languages.
- Human oversight and domain-specific AI training are vital for safe application in pediatric pharmacotherapy.
More Related Videos
05:47Evidence-based Knowledge Synthesis and Hypothesis Validation: Navigating Biomedical Knowledge Bases via Explainable AI and Agentic Systems
Published on: June 13, 2025
12:18A Machine Learning Approach to Design an Efficient Selective Screening of Mild Cognitive Impairment
Published on: January 11, 2020
Related Concept Videos
Improving Translational Accuracy
Improving Translational Accuracy