arXiv:2608.21415v1 Announce Type: cross
Abstract: Large Vision-Language Models (LVLMs) have achieved remarkable performance across a wide range of tasks; however, they often inherit social biases fro...
By Yisong Xiao, Aishan Liu, Yongxin Huang, Zonghao Ying, Shiji Zhao, Tianlin Li, Yong Han, Jian Yang, Xianglong Liu
arXiv:2509. 16462v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly used in high-stakes decision-making systems, where biased predictions can reinforce social and economic disparities.
By Mina Arzaghi, Alireza Dehghanpour Farashah, Florian Carichon, Jean-Fran\c{c}ois Plante, Golnoosh Farnadi
arXiv:2608.21577v1 Announce Type: cross
Abstract: Multimodal Large Language Models (MLLMs) are increasingly deployed in high-stakes domains where fairness is a critical safety requirement. In practic...
By Yuyang Luo, Kai Shu
FairLens is a benchmark and evaluation framework that measures fairness and validity of vision‑language models (VLMs) in high‑stakes domains such as hiring, legal, and healthcare. It uses over 100,000 face‑image and question pairs covering gender, race, and age, and assesses responses through demographic parity, soundness, demographic association, and bias in free‑text generation. The study finds that VLMs often make unwarranted inferences from faces rather than abstaining, especially in legal and healthcare contexts, and that small parity gaps can still hide unsafe treatment across groups.
By Vahid Reza Khazaie, Ahmed Y. Radwan, Shaina Raza
arXiv:2508.16748v2 Announce Type: replace-cross
Abstract: Prevalent multimodal self-supervised learning (SSL) methods rely on the redundancy assumption: that different views share substantial task-re...
By Jiaee Cheong, Abtin Mogharabin, Paul Liang, Hatice Gunes, Sinan Kalkan
arXiv:2510. 12857v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are now widely deployed in user-facing applications, reaching hundreds of millions of users worldwide.
By Robin Staab, Jasper Dekoninck, Maximilian Baader, Martin Vechev
The paper introduces FairTPT, a fairness-aware test‑time prompt tuning method for vision‑language models like CLIP. It jointly minimizes target marginal entropy while maximizing spurious marginal entropy to reduce bias under subpopulation shifts. Experiments show that standard episodic test‑time adaptation can worsen disparities, but FairTPT outperforms existing debiasing methods while preserving overall performance.
By Yoann Launay, Parameswaran Kamalaruban, Tom Kempton, Stuart Burrell, David Sutton
The paper introduces a method called Debias Anything that jointly addresses fairness and diversity in diffusion models without requiring sensitive-attribute annotations. By connecting a frozen diffusion model to a pretrained vision-language embedding space via an adapter, the approach uses pairs of text prompts to guide batch composition toward desired attribute proportions and employs a disagreement score to promote diversity. The method is applicable to both unconditional and text-conditional diffusion models and demonstrates improved quality and diversity while maintaining comparable fairness levels in experiments.
By Th\'eau d'Audiffret, Mariia Vladimirova, Jean-Yves Franceschi
arXiv:2604. 27011v2 Announce Type: replace-cross Abstract: AutoML, intended as the process of automating the application of machine learning to real-world problems, is a key step for AI popularisation.
By Alessia Berarducci, Eric Rossetto, Alessandro Antonucci, Marco Zaffalon
arXiv:2602. 06806v2 Announce Type: replace-cross Abstract: Text-to-image diffusion models achieve impressive generation quality but inherit and amplify training-data biases, skewing coverage of semantic attributes.
By Silpa Vadakkeeveetil Sreelatha, Dan Wang, Serge Belongie, Muhammad Awais, Anjan Dutta
arXiv:2404. 01356v3 Announce Type: replace-cross Abstract: Deep neural networks are vulnerable to adversarial perturbations that can simultaneously degrade prediction robustness and individual fairness across diverse application settings.
By Xuran Li, Hao Xue, Peng Wu, Xingjun Ma, Zhen Zhang, Huaming Chen, Flora D. Salim
arXiv:2606. 01282v1 Announce Type: cross Abstract: Text-to-Image (TTI) systems are now everyday infrastructure for journalism, education, advertising, and public communication, and the demographic and cultural stereotypes they inherit from training data (rendering women, people of colour, older adults, and non-Western cultures as under-represented or caricatured) become a population-level harm at deployment scale.
By Farbod Davoodi, Seyed Reza Tavakoli Shiyadeh, Pooria Safaei, Sana Harighi, Parsa Gholami, Amirali Amini, Kimia Vanaei, Emad Firoozi, Parham Abed Azad, Babak Khalaj, Siavash Ahmadi, Amir Hossein Payberah, Mohammad Hossein Rohban, Soheil Kolouri, Ali Diba