Comparing the Quality of Domain-Specific Versus General Language Models for Artificial Intelligence-Generated

Alireza Akhondi-Asl1,2,3,4, Youyang Yang1,2,3,4, Matthew Luchette1,2,3,4

  • 1Division of Critical Care Medicine, Department of Anesthesiology, Critical Care and Pain Medicine, Boston Children's Hospital, Boston, MA.

Summary

A smaller, domain-adapted language model (LM) fine-tuned on pediatric critical care notes outperformed larger general LMs in generating differential diagnoses. While still inferior to clinicians, these specialized LMs show potential as adjunct tools.