Classification performance bias between training and test sets in a limited mammography dataset

Rui Hou1,2, Joseph Y Lo2, Jeffrey R Marks3

  • 1Department of Artificial Intelligence, Beijing University of Posts and Telecommunications, Beijing, China.

Plos One
|February 7, 2024
PubMed
Summary

Data splitting in mammography radiomics studies can cause performance bias. Models trained on limited datasets may yield unreliable conclusions, highlighting the need for careful test set selection strategies.