arXiv AI By Julian Alfredo Mendez, Timotheus Kampik

The AR Fairness Metamodel: A Structured Framework for Fairness Measures

Read the original on arXiv AI →

The paper introduces the AR Fairness Metamodel, a structured framework for representing, analyzing, and comparing fairness scenarios. It incorporates key elements such as agents, resources, and their attributes, and supports both discrete and continuous fairness measures—including equality, equity, group fairness, individual fairness, the Gini index, the Theil index, Jain's fairness index, and a specific measure for Australia's Child Care Subsidy. The metamodel builds on the Tiles framework, offering modular components that can be connected to capture diverse fairness definitions, and includes formal proofs of relationships among group fairness, individual fairness, and envy‑freeness. An open‑source implementation of the Tiles framework is provided to facilitate practical fairness modeling and evaluation across various applications.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv AI
Aug 19

Position: Fairness Failure in Generative Models is an Evaluation Problem

The paper argues that fairness failures in generative models arise mainly from inadequate evaluation practices, making fairness findings hard to compare or use for deployment. It diagnoses common empirical and conceptual shortcomings in current methods and calls for a move toward standardized, generative‑specific evaluation. The authors introduce Fairness Cards, a minimal reporting artifact that explicitly documents evaluation choices—such as prompt families, counterfactual protocols, metrics, and refusal handling—to improve reproducibility, comparability, and accountability.

By Mariia Vladimirova, Jean-Yves Franceschi, Thibaut Issenhuth
arXiv Machine Learning
Sep 18

When fairness metrics fail: A utility-based perspective on $\varepsilon$-fairness

The paper argues that traditional probabilistic fairness metrics can miss significant disparities in the actual consequences of decisions. By introducing a utility-based framework, the authors show that a process can satisfy ε-fairness yet still be maximally unfair when utilities are considered. They apply this framework to college admissions and credit‑risk assessment, demonstrating that equalizing probabilities alone may mask unequal utility outcomes across groups.

By Tolulope Fadina, Thorsten Schmidt
arXiv Machine Learning
Sep 21

FairLMs: A Turnkey Library for Fairness in Language Models

FairLMs is a Python library designed to streamline fairness research in language models by unifying bias measurement, mitigation, and evaluation evidence. It offers 33 intrinsic and extrinsic metrics, 14 mitigation components across four intervention categories, 14 diagnostic tools, adapters for major Transformer architectures and hosted APIs, and benchmark loaders. The library enforces explicit declarations of model capabilities and input requirements, ensuring compatibility and reproducibility across components and datasets.

By Jiale Zhang, Michael Larionov, Zichong Wang, Zhipeng Yin, Wenbin Zhang