Two Centuries of Sexism in British Parliament: A Computational Analysis of Women's Representation in the Hansard Corpus
Read the original on arXiv Machine Learning →The Flow has not summarised this story yet — read it at arXiv Machine Learning.
The Flow has not summarised this story yet — read it at arXiv Machine Learning.
arXiv:2609.15207v1 Announce Type: new Abstract: Generative AI writing assistants and the Large Language Models (LLMs) that power them are increasingly part of how voters gather information before ele...
arXiv:2606.12186v2 Announce Type: replace Abstract: Enthymemes, arguments with unstated premises or conclusions, are pervasive in persuasive discourse, yet their annotation remains notoriously subjec...
arXiv:2608.30828v1 Announce Type: new Abstract: We present three large-scale studies of spoken parliamentary speech across four Slavic languages (Croatian, Czech, Polish, Serbian), drawing on over 6,...
Enthymemes, arguments with unstated premises or conclusions, are pervasive in persuasive discourse, yet their annotation remains notoriously subjective. We present a resource of 1,482 tweets from politically controversial discourse, annotated by five annotators for the presence of enthymemes and their argument structure, designed to study label variation.
The paper evaluates large language models (LLMs) for annotating German parliamentary debates on migration, achieving macro‑F1 scores comparable to human agreement, particularly with GPT‑5 and gpt‑oss‑120B. It combines soft‑label outputs with Design‑based Supervised Learning to mitigate systematic bias and applies the method to a 150‑plus‑year corpus, revealing high solidarity post‑war and a sharp rise in anti‑solidarity since 2015. The study demonstrates that LLMs can enable large‑scale social‑scientific analysis while highlighting the need for rigorous validation and bias correction.
arXiv:2608. 13410v1 Announce Type: new Abstract: Parliamentary proceedings are a primary record of democratic deliberation, yet their volume and fragmentation make multi-perspective access difficult for citizens, journalists, and researchers.