BRIDGE: benchmarking large language models for understanding real-world clinical practice texts

Jiageng Wu1, Bowen Gu1, Ren Zhou2

  • 1Division of Pharmacoepidemiology and Pharmacoeconomics, Department of Medicine, Brigham and Women's Hospital, Harvard Medical School, Boston, MA, USA.

Summary

A new benchmark, BRIDGE, evaluates large language models (LLMs) on real-world clinical data. It shows performance varies, with open-source LLMs matching proprietary ones and updated general models outperforming older specialized ones.

Related Concept Videos