arXiv AI By Xiang Yin, Tim Miller, Nico Potyka, Antonio Rago, Francesca Toni

Towards an Argumentative Foundation for Evaluative AI

Read the original on arXiv AI →

arXiv:2608. 07473v1 Announce Type: new Abstract: Evaluative AI (EAI) has been recently proposed as a way to support human decision-making, not by producing a single recommendation, but by presenting competing hypotheses together with evidence for and against each.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv AI
6d ago

OmouAI: Argumentative Human-AI Policy Deliberation with Simulated Personas

OmouAI is an interactive deliberation system that combines large language models with computational argumentation to facilitate policy debates involving humans and simulated personas such as stakeholders, experts, or devil’s advocates. Each persona generates its own arguments, which are assembled into a shared argumentation framework that users can contest, add to, or revise, ensuring human oversight. The system evaluates arguments using deterministic argumentative semantics against external goals like the UN Sustainable Development Goals, providing faithful explanations and indicating how policy recommendations affect those goals.

By Stylianos Loukas Vasileiou, Antonio Rago, William Yeoh, Georgina Curto
arXiv AI
Aug 20

A Theory of Post-hoc Debate Judgement

The paper proposes a theory for judging post-hoc debates in AI, focusing on properties like reproducibility, robustness, groundedness, and explainability. It evaluates two debate‑judgement methods—LLM judges and formal computational argumentation semantics—finding similar accuracy but noting that argumentation semantics offers stronger formal guarantees. The study suggests that argumentation semantics is a preferable framework for principled debate judges in AI systems.

By Xiang Yin, Adam Dejl, Antonio Rago, Lihu Chen, Francesca Toni
arXiv AI
Sep 3

Contrastive Explanations in Quantitative Bipolar Argumentation Frameworks

The paper introduces contrastive explanations for Quantitative Bipolar Argumentation Frameworks (QBAFs), a formalism used to represent and reason with information. Unlike traditional explanations that focus on a single argument, contrastive explanations highlight the differences between two topic arguments. The authors propose a general form of contrastive attribution functions (CAFs), present CAFs based on removal, gradients, and Shapley-values, and demonstrate their applicability in healthcare and bias identification contexts.

By Xiang Yin, Nico Potyka, Antonio Rago, Francesca Toni
arXiv AI
2d ago

ABDA-NL: A Natural-Language Scenario Explorer for Argument-Based Reasoning

ABDA-NL is a natural‑language interface for the ABDA argument‑based reasoning system, which uses ASPIC knowledge bases under grounded semantics. It lets users view accepted, rejected, or undecided conclusions, interactively explore the grounded discussion game, and experiment with what‑if scenarios by suspending assumptions, rules, or preferences. A large language model bridges natural language and formalism, answering questions from reference documents and translating plain‑English edits into formal statements, while the deterministic ABDA engine remains the sole source of arguments and acceptance labels, with all model proposals validated by the user before acceptance.

By Shawn Bowers, Martin Caminada, Haoyang Liu, Bertram Lud\"ascher
arXiv AI
Jul 23

Avoiding Obfuscation with Prover-Estimator Debate

arXiv:2506. 13609v2 Announce Type: replace Abstract: Training powerful AI systems to exhibit desired behaviors hinges on the ability to provide accurate human supervision on increasingly complex tasks.

By Jonah Brown-Cohen, Geoffrey Irving, Georgios Piliouras, Lijie Chen, Jiawei Li, Zhiyang Xun