Benchmarking the Confidence of Large Language Models in Answering Clinical Questions: Cross-Sectional Evaluation

Mahmud Omar1, Reem Agbareia2, Benjamin S Glicksberg1

  • 1Division of Data-Driven and Digital Medicine (D3M), Department of Medicine, Icahn School of Medicine at Mount Sinai, Gustave L. Levy Place New York, New York, NY, 10029, United States, 1 212 241 6500.

PubMed
Abstract