arXiv AI

Geopolitical Divisions Across Languages in Large Language Models

The study examines how language influences AI chatbot responses to questions about the war in Ukraine, revealing that the same AI systems (GPT, Claude, Gemini) produce varying political stances across 112 languages. By evaluating 20 statements in 112 languages, the researchers found that Russia‑leaning versus Ukraine‑leaning answers differ by language, mirroring global political attitudes such as public support for Russia, UN voting patterns, and aid levels. This pattern persists across all three models and even when specific statement pairs are removed, suggesting that information warfare could embed geopolitical biases into AI training data.

arXiv Machine Learning
Aug 20

Global Crises and National Policies: A Large Scale Analysis of Political Content in German Language Online Media

The study analyzes millions of German-language online articles and tweets from 2019–2022 to uncover political biases using automated text analysis. It finds that international events such as the COVID‑19 pandemic and the Ukraine war create thematic convergence between German and Swiss media, while domestic policy differences drive divergence in locally focused topics. Newspapers maintain more stable political content, whereas Twitter shows rapid, event‑driven spikes, illustrating how media platforms differ in intensity and timing.

By Yara D\"oring, Felix Bie{\ss}mann
arXiv AI
Aug 28

The BS-meter: Detecting Politics and Labour through ChatGPT's Language

The paper investigates the linguistic characteristics of ChatGPT-generated text, comparing it to 1,000 scientific publications and exploring its relation to concepts of ‘bullshit’ in political speech and workplace contexts. By applying hypothesis‑testing methods, the authors demonstrate that a statistical model of bullshit can link the artificial bullshit produced by ChatGPT to the political and workplace functions of bullshit observed in natural human language.

By Alessandro Trevisan, Harry Giddens, Sarah Dillon, Alan F. Blackwell
arXiv Computation and Language
Sep 14

SWARM: A Multilingual Human-Annotated Dataset for Russian Propaganda Detection in Search Engine Results

The paper introduces SWARM, a multilingual dataset of 2,183 search engine results in nine languages, annotated for support of Russian propaganda narratives. It evaluates a source-based blocklist, supervised classifiers, and zero‑shot large language models, finding that blocklists miss most propaganda and that content‑level models vary in performance, with the best LLM achieving an F1 of 0.73. The study highlights the need for per‑language, content‑level detection of search‑borne propaganda.

By Manuel Tonneau, Abhinav Dubey, Farhan Shaikh, Ilaria Vitulano, Martha Stolze, Hale Dedeoglu, Clara Riechert, Ella Kuka, Maryna Sydorova, Mykola Makhortykh, Elizaveta Kuznetsova
arXiv AI
Jul 21

Posts of Peril: Detecting Information About Hazards in Text

arXiv:2405. 17838v3 Announce Type: replace-cross Abstract: Socio-linguistic indicators of affectively-relevant phenomena, such as emotion or sentiment, are often extracted from text to better understand features of human-computer interactions, including on social media.

By Keith Burghardt, Daniel M. T. Fessler, Chyna Tang, Anne Pisor, Kristina Lerman