arXiv Machine Learning By Hadi Mohammadi, Tina Shahedi, Robert A. Bagheri, Mehdi Dastani, Masoume M. Raeissi

Learning Sexism Detection Using Multi-Agent Perspectivist Preference Optimization

Read the original on arXiv Machine Learning →

arXiv:2608. 04056v1 Announce Type: cross Abstract: When people label text for sexism, they often disagree, and not because some of them are wrong: they genuinely perceive sexism differently.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv AI
Sep 1

Evaluating and Mitigating Anti-LGBTQ Biases in German and Multilingual Language Models

The paper introduces a German-English benchmark dataset to evaluate anti‑LGBTQ biases in language models, combining community‑sourced stereotypes from German‑speaking queer individuals with a German translation of WinoQueer. Eight language models of varying sizes and architectures were assessed, revealing that they reproduce anti‑queer stereotypes with differences across identities and models. Fine‑tuning on community and progressive media content reduced bias on average, though the effect was not consistent across all models and identities.

By Melina Morch, Daniel Braun
arXiv AI
Jul 17

Step-Level Preference Learning for Generative Agents in Social Simulations

arXiv:2607. 14485v1 Announce Type: new Abstract: Large language model (LLM)-based generative agents simulate human behavior through long-horizon decision-making processes that comprise intermediate steps such as planning, memory retrieval, reflection, and action selection.

By Wenchang Gao, Pingyue Sheng, Lanlan Qiu, Yunfei Ma, Jian Zhao, Baicheng Chen, Kangda Wang, Yuyang Tian, Shunqiang Mao, Tianxing He