arXiv Computation and Language

FakeSpotter: A content and strategy agnostic Viral Misinformation Detection Tool

FakeSpotter is a new tool that estimates the viral misinformation risk of textual content by measuring structural fingerprints of misinformation instead of directly judging truthfulness. It operates across linguistic, narrative, logical, and critical‑thinking dimensions, using repeated large language model assessments and domain‑specific logistic regression classifiers for both short and long texts. In a labeled corpus of 764 texts, FakeSpotter achieved macro F1 scores of 0.788 for short texts and 0.793 for long texts, and its interpretive layer offers explainable outputs such as feature‑based scores, signal agreement, and a caution index for social listening.

arXiv Machine Learning
Sep 2

A Multi-Branch Feature Fusion Approach for Health Misinformation Detection and Propagation

This paper introduces a multi‑branch fusion framework that combines transformer‑based semantics, rhetorical cues, stance representations, and psychologically motivated proxies to detect health misinformation and characterize its spread on online social networks. The authors propose an interpretable Cognitive Propagation Score (CPS) derived from text cues that estimate argument complexity, emotional intensity, and virality potential, aiding diffusion‑risk reasoning when engagement data are missing. Experiments on three benchmark datasets (Constraint, COVID‑19_FNIR, Monkeypox) demonstrate near‑perfect classification and ranking performance, with ablation studies showing complementary gains from psychological and rhetorical components.

By Mkululi Sikosana, Sean Maudsley-Barton, Oluwaseun Ajao
arXiv AI
Aug 11

Build it, Break it, Repeat: Benchmarking and improving LLM-manipulated disinformation detection in social media posts

arXiv:2608. 09510v1 Announce Type: cross Abstract: Detecting machine-generated disinformation on social media is increasingly difficult as large language models (LLMs) make it easier to generate and rewrite misleading content at scale.

By Kevin Thomas, Milosz Kasprzyk, Reuel C Igbokwe Onuigbo, Elliott Pert, Cameron Tovey, Jo\~ao A. Leite, Olesya Razuvayevskaya, Carolina Scarton
arXiv Machine Learning
Sep 4

BharatGather: A Culturally-Informed Benchmark Dataset for Misinformation and Fake News Detection in Indian Public Events

BharatGather is a curated, multi-source dataset designed for binary misinformation classification in Indian public events such as religious festivals, political rallies, and cultural gatherings. The corpus contains 14,646 records assembled through systematic web scraping of fact‑checking platforms, multimedia transcript extraction, and LLM‑mediated synthetic augmentation to capture narrative diversity. It serves as a culturally informed benchmark to evaluate and develop fake‑news detection systems tailored to the socio‑cultural nuances of India’s mass‑gathering context.

By Parth Bramhecha, Smit Deshmukh, Sairaj Bodhale, Adwait Borate, Raviraj Joshi
arXiv AI
Aug 18

Propaganda Forensics: Recovering the Generation Pipeline of an AI-Driven Influence Campaign

arXiv:2608. 15746v1 Announce Type: new Abstract: We present a forensic analysis of the generation pipeline behind a recent AI-driven influence campaign.

By Benjamin Icard, Elouan Vuichard, Louis Lefebvre, Lila Sainero, Thomas Girault, Alice Breton, Tanguy Launay, Gauvain Bourgne, Morgane Casanova, Guillaume Gadek, Victor Kl\"otzer, Michel Le Nouy, Guillaume Gravier, Jean-Gabriel Ganascia, Paul \'Egr\'e