Related Experiment Videos

Benchmarking 54 large language model configurations for CAD-RADS scoring: open-weight models approach human-level

Veit Sandfort1,2, Davis M Vigneault3, Martin J Willemink4

  • 1Department of Radiology, Stanford University School of Medicine, 300 Pasteur Drive, Stanford, 94305, California, USA. veit.sandfort@gmail.com.

Summary

Large language models (LLMs) can now classify Coronary Artery Disease Reporting and Data System (CAD-RADS) categories in coronary CT angiography reports. An open-weight LLM achieved performance comparable to proprietary models, even on consumer hardware.