The study analyzes oral political language in U.S. presidential debates from 1960 to 2024, focusing on 19 candidates. It finds a clear trend toward simplification: sentence length and complex terms have decreased, while emotional tone has risen and logical, rational content has diminished. The research also explores whether specific presidents exhibit unique stylistic traits and whether language patterns correlate with electoral success.
By Jacques Savoy
The paper investigates the linguistic characteristics of ChatGPT-generated text, comparing it to 1,000 scientific publications and exploring its relation to concepts of ‘bullshit’ in political speech and workplace contexts. By applying hypothesis‑testing methods, the authors demonstrate that a statistical model of bullshit can link the artificial bullshit produced by ChatGPT to the political and workplace functions of bullshit observed in natural human language.
By Alessandro Trevisan, Harry Giddens, Sarah Dillon, Alan F. Blackwell
arXiv:2606. 28335v1 Announce Type: cross Abstract: We argue, with systematic empirical evidence, that a large language model's political ideology is not a fixed point, but a conditional distribution $\mathbb{P}($position$\mid$context$)$ over a real political space.
By Adib Sakhawat, Syed Rifat Raiyan, Tahsin Islam, Takia Farhin, Hasan Mahmud, Md Kamrul Hasan
arXiv:2512. 15792v4 Announce Type: replace-cross Abstract: Large language models (LLMs) have rapidly become indispensable tools for acquiring information and supporting human decision-making.
By Xulang Zhang, Rui Mao, Erik Cambria
arXiv:2512. 15792v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have rapidly become indispensable tools for acquiring information and supporting human decision-making.
By Xulang Zhang, Rui Mao, Erik Cambria
arXiv:2608.15641v2 Announce Type: replace
Abstract: This paper evaluates Wiktionary as an ethically crowdsourced lexicon for English dialects. We took a two-phase approach, providing an in-depth desc...
By Sidney Wong
arXiv:2509.22367v3 Announce Type: replace
Abstract: Large language models (LLMs) reflect politically-slanted opinions in their generated text. Even though it is widely assumed that model behavior ste...
By Tanise Ceron, Dmitry Nikolaev, Dominik Stammbach, Debora Nozza
arXiv:2606. 23462v2 Announce Type: replace-cross Abstract: Scientists do not, by profession, wage war.
By Sovesh Mohapatra, David Lydon-Staley, Dani S. Bassett
arXiv:2608. 05155v1 Announce Type: cross Abstract: Traditional sentiment analysis (SA) models, while effective for polarity classification, provide limited insight into the rhetorical, ideological, and framing dimensions of political discourse -- dimensions that are central to research in the social sciences and humanities (SSH).
By Maryam Fooladi, Federico Bottino
KoNeoBench is a curated dataset designed to evaluate large language models’ understanding of Korean neologisms. It contains 1,785 recently attested Korean words from online news since 2020, each accompanied by usage examples, word‑formation analyses, and dictionary‑style definitions. The authors define four evaluation tasks, report results from recent models and a human baseline, and find that current LLMs struggle with recovering source components, distinguishing semantic categories, and generating accurate definitions.
By Soha Lee, Soojin Lee, Heesung Yang, Hyunju Song, Hyunji Lee, Jinsan An, Jeongwan Shin, Jin Hyun Park, Jun Lee, Hyeyoung Park, Kilim Nam
PolERo presents a new dataset of 3,574 Romanian question‑answer pairs from presidential transcripts, annotated for political evasion using a two‑level taxonomy of response clarity and fine‑grained evasion strategies. The study evaluates various classification methods—including TF‑IDF baselines, fine‑tuned encoders, a sliding‑window encoder, and zero/few‑shot LLM prompting—under matched conditions. Cross‑lingual transfer experiments via joint bilingual training and machine‑translation augmentation reveal that fine‑tuned encoders perform competitively, transfer is asymmetric, and ambivalent evasion categories with pragmatic cues remain the most challenging across all models.
By Gabriel Stefan, Sergiu Nisioi