arXiv:2512. 15792v4 Announce Type: replace-cross Abstract: Large language models (LLMs) have rapidly become indispensable tools for acquiring information and supporting human decision-making.
By Xulang Zhang, Rui Mao, Erik Cambria
arXiv:2608. 06123v1 Announce Type: new Abstract: Measuring political bias in large language models (LLMs) remains challenging as it can manifest through subtle differences in framing, argumentation, and legal reasoning that are difficult to capture with a single metric.
By Massi-Nissa Abboud, Aladin Djuhera, Elena Cabrio, Holger Boche
arXiv:2509.22367v3 Announce Type: replace
Abstract: Large language models (LLMs) reflect politically-slanted opinions in their generated text. Even though it is widely assumed that model behavior ste...
By Tanise Ceron, Dmitry Nikolaev, Dominik Stammbach, Debora Nozza
arXiv:2607. 27232v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly shaping how we consume information and form our worldview.
By Haran Shani-Narkiss, Michael Fire, Oren Tsur
The study examines how language influences AI chatbot responses to questions about the war in Ukraine, revealing that the same AI systems (GPT, Claude, Gemini) produce varying political stances across 112 languages. By evaluating 20 statements in 112 languages, the researchers found that Russia‑leaning versus Ukraine‑leaning answers differ by language, mirroring global political attitudes such as public support for Russia, UN voting patterns, and aid levels. This pattern persists across all three models and even when specific statement pairs are removed, suggesting that information warfare could embed geopolitical biases into AI training data.
By Maxim Chupilkin
The paper introduces GPTBIAS, a framework that uses powerful large language models like GPT‑4 to evaluate bias in other LLMs. It employs specially crafted prompts called Bias Attack Instructions to probe for bias and outputs a bias score along with detailed information such as bias types, affected demographics, keywords, reasons, and improvement suggestions. Extensive experiments demonstrate the framework’s effectiveness and usability.
By Jiaxu Zhao, Meng Fang, Shirui Pan, Wenpeng Yin, Mykola Pechenizkiy