arXiv Computation and Language

Argument Structure Prediction in Online Conversations: A Comparative Study of Modeling Paradigms and Task Architectures

arXiv Computation and Language
Sep 23

Designing and Analysing Argument Mining Pipelines: Towards a Comprehensive Assessment

The paper introduces a meta‑study that reviews state‑of‑the‑art end‑to‑end argument mining (AM) pipelines. It proposes a triple‑perspective framework—linguistic, computational, and domain—to analyze how these pipelines model, compute, and incorporate domain knowledge into argument structures. The authors also outline a general design for the linguistic and computational aspects, aiming to standardize methodology descriptions and enable clearer comparisons among AM approaches.

By Siddharth Bhargava, Sara Tonelli, Patricia Mart\'in-Rodilla
arXiv AI
Jun 17

RooseBERT: A New Deal For Political Language Modelling

arXiv:2508. 03250v4 Announce Type: replace-cross Abstract: The increasing amount of political debates and politics-related discussions calls for the definition of novel computational methods to automatically analyse such content with the final goal of lightening up political deliberation to citizens.

By Deborah Dore, Elena Cabrio, Serena Villata
arXiv Computation and Language
Sep 10

Who Argues What? Joint Argument-Entity Detection and Classification in Political Debates

The paper introduces DNE‑ElecDeb, an enriched version of the USElecDeb dataset that annotates Debate Named Entities (DNEs) in both argumentative and non‑argumentative spans, and defines Debate Named Entity Recognition (DNER) as a new task. It proposes Joint Argument and Entity Tagging (JAET), a generative framework that fine‑tunes decoder‑only LLMs to insert inline argument and entity tags into debate turns while preserving the original transcript. JAET achieves significant improvements in joint AM+DNER performance (+27.3% relative F1 in the untyped setting and +41.9% in the typed setting) over sequential pipelines, and these gains generalize to Persuasive Essays (+26.6% and +52.7%).

By Lucio La Cava, Stefano Francesco Monea, Sergio Greco
arXiv AI
Aug 28

When Text Misleads: Inconsistent-Aware Reasoning for Audio-Grounded Dialogue

The paper introduces ContraTalk, a benchmark that tests whether dialogue models truly use acoustic cues or rely on transcript shortcuts. It formalizes cross‑modal disagreement, creates conflict and consistent QA examples, and proposes an Audio Twin representation to expose acoustic evidence to models. Experiments show that while text‑only LLMs perform well on consistent cases, they falter on conflict cases, and AudioLLMs only partially mitigate this issue.

By Yen-Ju Lu, Yuzhe Wang, Yaohan Guan, Xiluo He, Jiarui Hai, Mingrui Liang, Kaavya Chaparala, Thomas Thebaud, Laureano Moro-Velazquez, Najim Dehak, Jesus Villalba
arXiv Computation and Language
Sep 22

Extracting Arguments, Not Just Classifying Them: Instruction-Tuned LLMs for Generative Component Detection

arXiv:2609.24855v1 Announce Type: cross Abstract: Argumentative component detection (ACD) is a core subtask of Argument(ation) Mining (AM) and one of its most challenging aspects, as it requires join...

By Sofiane Elguendouze (UniCA, I3S, MARIANNE), Erwan Hain (UniCA, MARIANNE), Elena Cabrio (MARIANNE, UniCA), Serena Villata (CNRS, MARIANNE)
arXiv AI
3d ago

TTLab at Daleel 2026: STAR-Ar, Sequence Tagging for Argument Recognition in Arabic

The paper introduces STAR‑Ar, a BERT‑BiLSTM‑CRF model designed for the Daleel 2026 Arabic argument mining shared task. It treats argument discourse unit detection and classification as a token‑level sequence labeling problem, achieving an F1‑score of 72.69 on validation and 73.7 on test data. Analysis shows that models trained only on editorial texts perform worse than those trained on debates, mainly due to the smaller editorial dataset.

By Bhuvanesh Verma, Ali Abusaleh, Alexander Mehler