arXiv AI

Position: Beyond Sensitive Attributes, ML Fairness Should Quantify Structural Injustice via Social Determinants

arXiv:2508. 08337v3 Announce Type: replace-cross Abstract: Algorithmic fairness research has largely framed unfairness as discrimination along sensitive attributes.

arXiv Machine Learning
Sep 15

From Network Inequality to Network Fairness: A Perspective on Responsible Decision-Making

The article discusses how social networks influence decision-making and opportunity distribution, noting that network-generating mechanisms often mirror existing inequalities and can amplify disparities when used in technology. It identifies ten network effects that bias the link between intended measurements and observed data, using academic hiring as a case study to show that network biases can be neither inherently harmful nor beneficial. The authors argue for a comprehensive, networked fairness framework that incorporates both distributive and procedural justice and involves all stakeholders.

By Lisette Esp\'in-Noboa, Tina Eliassi-Rad, Pak-Hang Wong, Erich Prem, Meike Zehlike, Ricardo Baeza-Yates, Suresh Venkatasubramanian, Fariba Karimi
arXiv Machine Learning
Sep 18

When fairness metrics fail: A utility-based perspective on $\varepsilon$-fairness

The paper argues that traditional probabilistic fairness metrics can miss significant disparities in the actual consequences of decisions. By introducing a utility-based framework, the authors show that a process can satisfy ε-fairness yet still be maximally unfair when utilities are considered. They apply this framework to college admissions and credit‑risk assessment, demonstrating that equalizing probabilities alone may mask unequal utility outcomes across groups.

By Tolulope Fadina, Thorsten Schmidt
arXiv AI
Jul 13

Tuning Derivatives for Causal Fairness in Machine Learning

arXiv:2605. 05882v2 Announce Type: replace-cross Abstract: Artificial-intelligence systems are becoming ubiquitous in society, yet their predictions typically inherit biases with respect to protected attributes such as race, gender, or age.

By Filip Edstr\"om, Guilherme W. F. Barros, Tetiana Gorbach, Xavier de Luna
arXiv Machine Learning
Aug 19

Advancing Health Equity through Multi-Level Fairness in Health Informatics

The paper examines how multi‑level fairness techniques—combining several bias‑mitigation steps—can reduce biases across patient demographics in health informatics. It reviews current literature, identifies gaps in implementation and reporting of health equity outcomes, and evaluates the role of reporting standards such as MINIMAR and TRIPOD in enhancing transparency. The authors conclude with recommendations to improve reporting transparency, broaden adoption of multi‑level fairness methods, and explicitly prioritize health equity in future research.

By Nick Souligne, Vignesh Subbian
arXiv Machine Learning
Sep 24

When Post-Processing Fairness Constraints Help and When They Harm: Evidence from Eight Cross-Domain Evaluations

The paper introduces FAPE, a four‑stage framework for evaluating the post‑processing fairness intervention ThresholdOptimizer across eight diverse domains, including criminal justice, finance, healthcare, and education. It reports that the intervention reduces disparity in most high‑disparity cases but can worsen fairness when baseline disparities are low, and that a single deployment‑time audit is unreliable without continuous monitoring and baseline‑disparity screening.

By Nithin Raghava Ramachandra Narla
arXiv AI
Sep 10

PopResume: Causal Fairness Evaluation of LLM/VLM Resume Screeners with Population-Representative Dataset

PopResume is a population‑representative resume dataset designed for causal fairness auditing of large language model (LLM) and vision‑language model (VLM) resume screeners. It grounds fairness evaluation in real population statistics and preserves natural attribute relationships, enabling path‑specific effect (PSE) analysis that separates business‑necessity from redlining pathways. Using PopResume, the authors evaluated eight models on 60.8K resumes across five occupations and uncovered five discrimination patterns that aggregate metrics missed, demonstrating the value of causally‑grounded auditing.

By Sumin Yu, Juhyeon Park, Taesup Moon