Related Experiment Video
Updated: Jul 4, 2025

Author Spotlight: Biological Standardization to Ensure Reproducibility and Harmonization in Research
Published on: August 4, 2023
Improving the Reliability of Peer Review Without a Gold Standard
Tarmo Äijö1, Daniel Elgort2,3, Murray Becker2,4
1Covera Health, New York, NY, USA. tarmo.aijo@coverahealth.com.
None:
Peer review plays a crucial role in accreditation and credentialing processes as it can identify outliers and foster a peer learning approach, facilitating error analysis and knowledge sharing. However, traditional peer review methods may fall short in effectively addressing the interpretive variability among reviewing and primary reading radiologists, hindering scalability and effectiveness. Reducing this variability is key to enhancing the reliability of results and instilling confidence in the review process. In this paper, we propose a novel statistical approach called "Bayesian Inter-Reviewer Agreement Rate" (BIRAR) that integrates radiologist variability. By doing so, BIRAR aims to enhance the accuracy and consistency of peer review assessments, providing physicians involved in quality improvement and peer learning programs with valuable and reliable insights. A computer simulation was designed to assign predefined interpretive error rates to hypothetical interpreting and peer-reviewing radiologists. The Monte Carlo simulation then sampled (100 samples per experiment) the data that would be generated by peer reviews. The performances of BIRAR and four other peer review methods for measuring interpretive error rates were then evaluated, including a method that uses a gold standard diagnosis. Application of the BIRAR method resulted in 93% and 79% higher relative accuracy and 43% and 66% lower relative variability, compared to "Single/Standard" and "Majority Panel" peer review methods, respectively. Accuracy was defined by the median difference of Monte Carlo simulations between measured and pre-defined "actual" interpretive error rates. Variability was defined by the 95% CI around the median difference of Monte Carlo simulations between measured and pre-defined "actual" interpretive error rates. BIRAR is a practical and scalable peer review method that produces more accurate and less variable assessments of interpretive quality by accounting for variability within the group's radiologists, implicitly applying a standard derived from the level of consensus within the group across various types of interpretive findings.
More Related Videos
08:29An Open Source Technology Platform to Manufacture Hydrogel-Based 3D Culture Models in an Automated and Standardized Fashion
Published on: March 31, 2022
08:52Improving Reproducibility to Meet Minimal Information for Studies of Extracellular Vesicles 2018 Guidelines in Nanoparticle Tracking Analysis
Published on: November 17, 2021
Related Concept Videos
Improving Translational Accuracy
Reliability and Validity
Blind Procedures
Self-Evaluation: Self-Enhancement and Self-Verification
Data Validation
Key parameters for method validation include:
Ethics in Research