Related Experiment Video
Updated: Jan 8, 2026

Author Spotlight: Exploring Sex-Specific Glial Signatures and Therapeutic Leads for Alzheimer's Disease
Published on: May 20, 2024
Effectiveness of ChatGPT, Google Gemini, and Microsoft Copilot in Answering Thai Drug Information Queries:
Suphannika Pornwattanakavee1, Nattawut Leelakanok1, Teerarat Todsarot1
1Division of Clinical Pharmacy, Faculty of Pharmaceutical Sciences, Burapha University, Chonburi, Thailand.
Background:
ChatGPT-4o, Google Gemini, and Microsoft Copilot have shown potential in generating health care-related information. However, their accuracy, completeness, and safety for providing drug-related information in Thai contexts remain underexplored.
Objective:
This study aims to evaluate the performance of artificial intelligence (AI) systems in responding to drug-related questions in Thai.
Methods:
An analytical cross-sectional study was conducted using 76 public drug-related questions compiled from medical databases and social media between November 1, 2019, and December 31, 2024. All questions were categorized into 19 distinct categories, each comprising 4 questions. ChatGPT-4o, Google Gemini, and Microsoft Copilot were queried in a single session on March 1, 2025, by using input in Thai. All responses were evaluated for correctness, completeness, risk, and reproducibility independently by clinical pharmacists using standardized evaluation criteria.
Results:
All 3 AI models provided generally complete responses (P=.08). ChatGPT-4o yielded the highest proportion of fully correct responses (P=.08). The overall risk levels of high-risk answers were not significantly different (P=.12). Response correctness was influenced by the category of the drug-related questions (P=.002) but not completeness (P=.23). The correctness of Google Gemini and Microsoft Copilot was higher than that of ChatGPT for pharmacology queries. The type of questions also statistically significantly affected the risk level of the answers (P=.04). In particular, the pregnancy and lactation category had the highest high-risk response rate (1/76, 1% per system). All 3 AI models demonstrated consistent response patterns when the same questions were re-queried after 1, 7, and 14 days.
Conclusions:
The evaluated AI chatbots were able to answer the queries with generally complete content; however, we found limited accuracy and occasional high-risk errors in responding to drug-related questions in Thai. All models exhibited good reproducibility.
More Related Videos
05:47Evidence-based Knowledge Synthesis and Hypothesis Validation: Navigating Biomedical Knowledge Bases via Explainable AI and Agentic Systems
Published on: June 13, 2025
08:15Author Spotlight: Network Pharmacology and Molecular Docking to Decipher the Action of Jiawei Shengjiang San Against Diabetic Kidney Disease
Published on: May 10, 2024
Related Concept Videos
Effect of Hepatic Disease on Pharmacokinetics: Dose Adjustments Due to Hepatic Impairment
Drug Dosing: Geriatric Patients
Therapeutic Drug Monitoring: Drug Analysis Methods
Factors Affecting Drug Response: Overview
Drug Dosage Regimen: Overview
Typically, the starting dose and dosing interval are guided by the manufacturer's recommendations based on clinical trials conducted during and after drug...
Therapeutic Drug Monitoring: Affecting Factors