Leveraging large language models for structured information extraction from pathology reports

Jeya Balaji Balasubramanian1, Daniel Adams2,3, Ioannis Roxanis4

  • 1Division of Cancer Epidemiology and Genetics, National Cancer Institute, 9609 Medical Center Dr, NCI Shady Grove, Room 7E554, Rockville, MD 20850, USA.

PubMed
Summary

Large language models (LLMs) achieve human-level accuracy in extracting structured data from breast cancer histopathology reports. This automated approach enhances data accessibility for clinical research, offering a scalable alternative to manual extraction.