Validating Radiology Artificial Intelligence Model Performance on Photon-Counting CT Images Using Large Language

Yee Seng Ng1, Mohammed M Kanani2, William E King1

  • 1Department of Radiology, University of Washington, Seattle, Washington.

Summary

Large language models (LLMs) automate ground truth label extraction from radiology reports, enabling scalable assessment of artificial intelligence (AI) tools. This method reliably validates AI performance, even with new imaging hardware like photon-counting CT scanners.