arXiv AI By Dahyun Lee, Jiyoung Han, Kunwoo Park

Visual Framing for News Stance Detection via Image Generation

Read the original on arXiv AI →

The paper introduces VFStance, a method that uses image generation to make implicit stance cues in news articles more explicit through visual framing. It targets article-level news stance detection, a task complicated by subtle, structurally complex texts. Experiments show VFStance outperforms existing methods, and a user study with 200 participants demonstrates that the visual framing makes stance signals more noticeable in a snippet-based news consumption setting.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv Computation and Language
3d ago

MMDS-Bench: Benchmarking Multimodal Large Language Models on Dynamic Stance in Social Media Interactions

MMDS-Bench is a new diagnostic benchmark for multimodal dynamic stance classification in social media parent‑reply interactions. It contains 3,482 multimodal instances annotated with a seven‑label stance taxonomy, plus an 800‑instance subset that demands structured reasoning over parent and reply understanding and stance‑relation inference. The benchmark also tags each instance with five challenge factors—multimodal fusion, parent framing, non‑literal expression, interaction reasoning, and label‑boundary ambiguity—and evaluates 12 multimodal large language models using a reference‑grounded LLM‑judge protocol, revealing that current models still struggle with relational inference beyond separate parent and reply comprehension.

By Yuzhe Ding, Kang He, Li Zheng, Shengwu Zheng, Teng Shi, Fei Li, Chong Teng, Donghong Ji