Related Experiment Video
Updated: Dec 8, 2025

09:47
Spotting Cheetahs: Identifying Individuals by Their Footprints
Published on: May 1, 2016
15.2K
Forensic Footwear Reliability: Part III-Positive Predictive Value, Error Rates, and Inter-Rater Reliability
Nicole Richetelli1, Lesley Hammer2, Jacqueline A Speir1
1West Virginia University, 208 Oglebay Hall, PO Box 6121, Morgantown, WV, 26506.
Journal of Forensic Sciences
|September 22, 2020
Summary
Forensic footwear examiners in the US show substantial agreement, with accuracy metrics like positive predictive value at 98.8% and negative predictive value at 93.3% in this extensive study.
Area of Science:
- Forensic Science
- Criminalistics
- Pattern Evidence Analysis
Background:
- Forensic footwear examination is crucial for linking suspects to crime scenes.
- Estimating examiner performance and reliability is vital for courtroom admissibility.
- Previous studies have varied in methodology and scope, necessitating further research.
Purpose of the Study:
- To quantify the performance of forensic footwear examiners in the United States.
- To determine error rates, predictive values (PV), and inter-rater reliability (IRR).
- To evaluate the impact of different reporting structures on examiner agreement.
Main Methods:
- Collected data from 70 footwear experts over 19 months, involving 12 comparisons each.
- Analyzed over 1000 examiner attributes, 3500 impression features, and 840 source conclusions.
- Calculated error rates, PVs, and Gwet AC2 agreement coefficients for examiner reproducibility.
Main Results:
- Correct predictive value ranged from 94.5% (exclusions) to 85.0% (identifications).
- False-positive rate was 0.48%, false-negative rate was 15.6%.
- Gwet AC2 agreement coefficients indicated substantial (0.751) to moderate (0.692) agreement between raters.
Conclusions:
- Forensic footwear examination demonstrates acceptable performance metrics and reliability.
- A six-level reporting structure shows higher inter-rater reliability compared to a four-level structure.
- The findings contribute to understanding the scientific validity of footwear evidence analysis.
Related Concept Videos
Reliability and Validity
13.6K
Reliability and validity are two important considerations that must be made with any type of data collection. Reliability refers to the ability to consistently produce a given result. In the context of psychological research, this would mean that any instruments or tools used to collect data do so in consistent, reproducible ways.
13.6K
Sensitivity, Specificity, and Predicted Value
1.1K
In healthcare diagnostics, laboratory tests play a crucial role in identifying and diagnosing a wide range of medical conditions. However, interpreting test results is not always straightforward. An abnormal test result does not always confirm the presence of a disease, just as a normal result does not guarantee its absence. To assess the reliability of these diagnostic tools, healthcare practitioners rely on two key statistical indicators: sensitivity and specificity.
Sensitivity is the...
Sensitivity is the...
1.1K
Accuracy and Errors in Hypothesis Testing
481
Hypothesis testing is a fundamental statistical tool that begins with the assumption that the null hypothesis H0 is true. During this process, two types of errors can occur: Type I and Type II. A Type I error refers to the incorrect rejection of a true null hypothesis, while a Type II error involves the failure to reject a false null hypothesis.
In hypothesis testing, the probability of making a Type I error, denoted as α, is commonly set at 0.05. This significance level indicates a 5%...
In hypothesis testing, the probability of making a Type I error, denoted as α, is commonly set at 0.05. This significance level indicates a 5%...
481
Bonferroni Test
3.2K
The Bonferroni test is a statistical test named after Carlo Emilio Bonferroni, an Italian mathematician best known for Bonferroni inequalities. This statistical test is a type of multiple comparison test to determine which means are different than the rest. Bonferroni test can minimize the Type 1 error by reducing the significance level alpha, which otherwise increases with sample pairs.
The means of different samples are first paired in all possible combinations.
The null hypothesis of the...
The means of different samples are first paired in all possible combinations.
The null hypothesis of the...
3.2K

