Related Experiment Video
Updated: May 24, 2026

04:57
Assisted Selection of Biomarkers by Linear Discriminant Analysis Effect Size (LEfSe) in Microbiome Data
Published on: May 16, 2022
DeepSeek R1 Distilled Fails to Perform Well Against the USMLE and Other LLMs with and Without Semantics
Peter L Elkin1, Guresh Mehta1, Aaron N Elkin1
1Department of Biomedical Informatics, University at Buffalo, USA.
Studies in Health Technology and Informatics
|May 23, 2026
Abstract:
DeepSeek-R1 Distilled Large Language model (LLM) was touted to be almost as good and much cheeper to generate than traditional LLMs. We compared its ability to answer medical questions from the United States Medical Licensing Examinations (USMLE) and found that the performance drop from the original model to DeepSeek was significant. DeepSeek is not yet a good alternative for medical question answering.
