arXiv AI By Jiale Dai, Liuxian Ma, Xiaoke Niu, Wenjing Zhang, Huiying Zhao, Zhaoxiang Liu, Shiguo Lian, Guojie Song

VISTA: Value-Informed Event Appraisal for Multimodal Emotion Conflict

Read the original on arXiv AI →

VISTA (Value-Informed Semantic Trust Arbitration) is a learned seven-field appraisal interface that conditions modality arbitration on concerns, event relations, and expression conditions while retaining a joint-evidence residual. It uses a log-odds decomposition to separate emotion expectation from cue diagnosticity, allowing appraisal to change how evidence is interpreted. With a shared Qwen2.5-Omni-7B backbone, VISTA achieves 64.5% conflict accuracy on CA-MER, improving on modality gating by 2.5 percentage points on conflict and 0.2 on consistency, and a frozen-backbone probe reaches 0.600 macro CCC for appraisal readout versus 0.505 for emotion-only fine-tuning.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv Computation and Language
Sep 4

Chiaroscuro for Emotions: A Contrastive Emotion Benchmark Grounded in Appraisal Theory

The paper introduces CHIARO, a 1,000-sentence benchmark for contrastive emotion inference grounded in appraisal theory, where each scenario elicits a positive emotion in one person and a negative emotion in another. The dataset covers ten emotion classes and is human‑annotated. Evaluation shows that the best large language model achieves 67.3 macro‑F1, below human agreement, while existing emotion classifiers perform near chance. When used as a training signal alongside an existing emotion corpus, models improve on CHIARO and on six of ten external emotion benchmarks, demonstrating its value as a complementary training resource.

By Divyesh Bommana, Mohammad Saim, Tianyu Jiang
arXiv AI
Sep 10

EmoMed: An Emotionally-Aware Agent for Multimodal Medical Support with Real-Time Information Retrieval

EmoMed is a multimodal medical consultation agent that tailors its responses to users' emotional states—such as anxiety, confusion, or urgency—while preserving clinical accuracy. It processes text and medical images, detects affect indicators, and adjusts tone, structure, and detail accordingly. The system ensures factual reliability through a dual retrieval mechanism that combines web-based fact‑checking with an API‑connected, continuously updated medical knowledge base, and it has been evaluated across seven state‑of‑the‑art language models using comprehensive metrics, showing that emotionally adaptive responses outperform neutral baselines without sacrificing accuracy.

By Ivan Nasonov, Nikita Glazkov, Ivan Makovetskiy, Mikhail Mozikov, Daniil Sukhorukov, Andrey Savchenko, Ilya Makarov