Related Experiment Video
Updated: Feb 4, 2026

Author Spotlight: Addressing Technical and Subjective Challenges in Measuring Classroom Attention
Published on: December 15, 2023
Discovery of an Artificial Intelligence Label Feedback Loop: How the Success of a Clinically Implemented Artificial
Patrick L Day1, Mikolaj A Wieczorek2, Denise Rokke1
1Department of Laboratory Medicine and Pathology, Mayo Clinic, Rochester, MN, United States.
Artificial intelligence (AI) in lab tests can create an "AI label feedback loop," negatively impacting model retraining and evaluation. Careful human oversight and validation are crucial for reliable AI system performance in clinical settings.
Area of Science:
- Medical Diagnostics
- Artificial Intelligence in Healthcare
- Laboratory Medicine
Background:
- Artificial intelligence (AI) enhances laboratory tests, improving quality and efficiency.
- Clinical AI implementation can lead to unforeseen challenges in model retraining.
- This study identifies an "AI label feedback loop" in an AI-augmented kidney stone composition test.
Purpose of the Study:
- To describe the discovery of an AI label feedback loop.
- To analyze its impact on model retraining and evaluation.
- To propose mitigation strategies for clinical AI systems.
Main Methods:
- Evaluated two versions (V1 and V2) of an AI-augmented kidney stone composition test.
- V2 utilized six times more data than the initial V1 model.
- Performance was assessed on three datasets: AI-influenced labels, pre-AI human labels, and human-corrected labels.
Main Results:
- V2 showed a 10% lower concordance than V1 on AI-influenced data, despite more training data.
- Performance was similar between V1 and V2 on pre-AI, human-labeled data.
- V2 outperformed V1 on human-corrected data, especially for rare kidney stone types, revealing the feedback loop.
Conclusions:
- AI integration in clinical practice can influence test results, complicating AI model development and evaluation.
- Ongoing human annotation and meticulous validation set construction are vital.
- These strategies ensure reliable AI performance assessment and support safe clinical AI evolution.
Related Concept Videos
Intelligence
Feedback Loops
Trial and Error and Algorithm
Measures of Intelligence
Validity refers to how well a test measures what it claims to measure. An intelligence test should accurately assess intelligence rather than another characteristic, like anxiety. Criterion validity is one way to evaluate this;...
Multiple Intelligences Theory
Cattell's Theory of Intelligence
Fluid intelligence involves the capacity to solve new problems and adapt to unfamiliar situations. It's the type of intelligence individuals use when they encounter a novel problem or puzzle that requires innovative thinking. For instance, figuring out how to operate a new gadget relies heavily on...

