Evaluating the Performance of AI Large Language Models in Detecting Pediatric Medication Errors Across Languages: A

Rana K Abu-Farha1, Haneen Abuzaid2, Jena Alalawneh3

  • 1Clinical Pharmacy and Therapeutics Department, Faculty of Pharmacy, Applied Science Private University, Amman 11937, Jordan.

PubMed
Summary

Microsoft Copilot showed the highest accuracy in detecting pediatric medication errors among four AI models. Performance varied by language, with Arabic generally showing lower accuracy, highlighting the need for better multilingual AI training.