GUIDE DE LA SOCIÉTÉ

The Impossibility Theorem of Fairness

Several impossibility results show that common statistical fairness criteria can conflict.

  • 3 minutes de lecture
  • Dernière mise à jour
Sur cette page3 minutes de lecture
  1. Aperçu
  2. Plongée profonde
  3. Impact stratégique
  4. The Future of The Impossibility Theorem of Fairness
  5. Mise en œuvre dans le monde réel
  6. Risques et garde-fous
  7. Feuille de route de mise en œuvre
  8. Continuez à explorer
  9. Questions fréquemment posées

Aperçu

For risk scores with different group base rates and imperfect prediction, calibration and two score-balance conditions—mean scores within positive and negative outcome groups—cannot generally all hold exactly; related classifier results concern thresholded error rates.

Plongée profonde

“The impossibility theorem” is shorthand for several results about incompatible statistical fairness criteria. Kleinberg, Mullainathan and Raghavan’s 2016 paper formalizes three score-level conditions: calibration within groups; balance for the positive class, meaning the average score among people with a positive outcome is equal across groups; and balance for the negative class, meaning the average score among people with a negative outcome is equal across groups. Except in constrained cases—including equal base rates or perfect prediction—one cannot satisfy all three simultaneously. These positive- and negative-class balance conditions compare average scores conditional on outcomes; they are not directly false-positive or false-negative rate parity after thresholding. Calibration means that among people assigned a given score, the observed outcome frequency matches that score within each group. A classifier applies a threshold to a score and assigns a discrete decision. Chouldechova’s related 2017 result addresses a thresholded classifier: with differing outcome prevalence and an imperfect predictor, predictive parity and equal false-positive/false-negative rates cannot both generally hold. That classifier-level error-rate balance is related to, but distinct from, the score-level positive- and negative-class balance conditions in Kleinberg et al. These results depend on the target, base rates, score quality and chosen criteria; they do not establish a universal impossibility of fairness. The theorem does not tell a decision maker which criterion to prioritize, establish that observed labels are valid ground truth, or guarantee that optimizing any one metric produces just outcomes. Context matters: false positives and false negatives may have different costs, scores may be used as rankings rather than probabilities, and group labels or outcomes can themselves reflect structural disadvantage. Woodworth and colleagues study conditions under which one can learn fair representations while navigating trade-offs, further illustrating that conclusions depend on the formal setup. Responsible use requires stating the assumptions and policy choice, reporting resulting harms, and considering procedural and causal approaches beyond the statistical parity metrics.

Impact stratégique

Risques et sécurité

Les dommages catastrophiques et quotidiens causés par l’IA dépendent tous deux de la personne qui comprend les risques et qui peut agir.

Décisions plus claires

Les connaissances du public et des professionnels déterminent si une politique de sécurité forte est politiquement possible.

Passer à travers le battage médiatique

Des explications claires réduisent la capture par le battage médiatique, les relations publiques en laboratoire et le théâtre d'éthique vague.

The Future of The Impossibility Theorem of Fairness

Fairness research continues to refine conditions for scores, classifiers, multiple groups and uncertain labels. Treat each theorem as scoped to its assumptions, and revisit metric choices when base rates, outcome definitions or decision thresholds change. Keep a dated record of the primary source or study behind each claim and revisit conclusions when new evidence or implementation details emerge. Newer work examines alternative score types, more than two groups and imperfect labels. These extensions do not remove the need to state assumptions and the actual policy objective.

Mise en œuvre dans le monde réel

A risk-score team states whether it prioritizes calibration, equalized error rates or another criterion before comparing groups.

A lending audit reports base rates and false-positive/false-negative rates so stakeholders can see which fairness properties conflict.

A policy maker explains why one metric is prioritized given the decision’s harms instead of claiming a mathematical theorem selected the policy.

A model reviewer tests whether a claimed trade-off applies because group base rates differ and predictions are imperfect.

Risques et garde-fous

  • Traiter le risque existentiel comme de la science-fiction alors que les capacités s’accroissent.

  • Confondre sécurité des produits de surface et alignement sous haute autonomie.

  • Laisser le public non anglophone et non expert avec uniquement des sources de mauvaise qualité.

Feuille de route de mise en œuvre

  1. Séparez les dommages causés aux produits, leur mauvaise utilisation et les risques de perte de contrôle/désalignement.

  2. Demandez quelles preuves pourraient changer votre point de vue sur les délais et la gravité.

  3. Préférez les sources primaires et les évaluations concrètes aux allégations marketing.

  4. Identifiez une voie d’action : carrière, politique, financement ou compétences – et pas seulement la sensibilisation.

Continuez à explorer

Free newsletter

Get the daily AI briefing

Three verified AI stories every weekday morning, written in plain English. Free forever, no ads.

One email each weekday. Unsubscribe in one click. We never sell or share your address.

Test yourself

Take the The Impossibility Theorem of Fairness quiz

Instant feedback on every answer, and a shareable certificate with a verifiable ID once you pass a course.

Démarrer le quiz

Support free AI education. AI Understanding is a 501(c)(3) nonprofit — no ads, no paywall, ever. Make a donation

Questions fréquemment posées

What is The Impossibility Theorem of Fairness?

Several impossibility results show that common statistical fairness criteria can conflict. For risk scores with different group base rates and imperfect prediction, calibration and two score-balance conditions—mean scores within positive and negative outcome groups—cannot generally all hold exactly; related classifier results concern thresholded error rates.

Which three conditions are central to the Kleinberg–Mullainathan–Raghavan result?

The paper formalizes calibration and balance conditions for positive and negative classes.

Under what common special case can the three conditions avoid the stated conflict?

The theorem identifies constrained cases including equal base rates or perfect prediction.

What does calibration within groups mean?

Calibration means the score corresponds to observed outcome frequencies within each group.

For a thresholded classifier, which rates does Chouldechova’s related error-rate-balance criterion compare?

This related classifier-level criterion compares false-positive and false-negative rates across groups after a decision threshold; it is distinct from KMR’s score-level balance conditions.

What role do differing base rates play in the incompatibility results?

The incompatibility arises under differing group prevalences when predictions are not perfect.