Related Experiment Video
Updated: Sep 18, 2025

Selecting Multiple Biomarker Subsets with Similarly Effective Binary Classification Performances
Published on: October 11, 2018
Building Nondiscriminatory Algorithms in Selected Data
David Arnold1, Will Dobbie2, Peter Hull3
1University of California, San Diego and NBER.
None:
We develop new quasi-experimental tools to understand algorithmic discrimination and build non-discriminatory algorithms when the outcome of interest is only selectively observed. We first show that algorithmic discrimination arises when the available algorithmic inputs are systematically different for individuals with the same objective potential outcomes. We then show how algorithmic discrimination can be eliminated by measuring and purging these conditional input disparities. Leveraging the quasi-random assignment of bail judges in New York City, we find that our new algorithms not only eliminate algorithmic discrimination but also generate more accurate predictions by correcting for the selective observability of misconduct outcomes.
Related Concept Videos
Stereotypes, Prejudice, and Discrimination
Mechanistic Models: Compartment Models in Algorithms for Numerical Problem Solving
In individual population analyses, different algorithms are employed, such as Cauchy's method, which uses a...
Stratified Sampling Method
To choose a stratified sample, divide the population into groups called strata and then take a...
Bias
In statistics, a sampling bias is created when a sample is collected from a population, and some members of the population are not as likely to be chosen as others (remember, each member...
Quantifying and Rejecting Outliers: The Grubbs Test
Statistical Inference Techniques in Hypothesis Testing: Parametric Versus Nonparametric Data
Parametric statistics, as the name suggests, assumes that data follow a specific distribution, often a normal distribution. This assumption enables robust hypothesis testing and estimation. Parametric methods, like the Student's t-test or Goodness-of-fit test, are frequently employed in biostatistics due to their robustness. For instance,...

