← Back to all news
arXiv Computation and Language September 16, 2026 By Leon Hammerla, Patrick Schrottenbacher, Alexander Mehler

Negation Beyond the Verbal Channel: Temporal Multimodal Correlates in Dialogue

Read the original on arXiv Computation and Language →

The Flow has not summarised this story yet — read it at arXiv Computation and Language.

  • multimodal
  • benchmarks

One email a morning, machine-written

One email a day, machine-written, one click to leave. We never share your address.

Related stories

arXiv Computation and Language
3d ago

Exploring Multimodal Turn-Taking Cues in Face-to-Face Conversation using Voice Activity Projection

arXiv:2609.14666v1 Announce Type: cross Abstract: Turn-taking is a fundamental component of spoken interaction, and while humans naturally rely on both verbal and non-verbal signals, dialogue systems...

By Willem Berner, Julio Cesar Cavalcanti, Kalle {\AA}str\"om, Gabriel Skantze
llmsmultimodal
More like this →
arXiv AI
Aug 25

LLM-Based Selection of Incongruent Verbal and Nonverbal Behavior for Virtual Humans

arXiv:2608.22731v1 Announce Type: new Abstract: Nonverbal behavior generation systems for virtual agents often take an utterance as input and generate nonverbal behaviors that emphasize or illustrate...

By Parisa Ghanad Torshizi, Stacy Marsella
llmsagentsrobotics
More like this →
arXiv Computer Vision
Aug 24

EmotionDialogCN: A Spontaneous Multimodal Dataset for Mandarin Emotional Dialogue

arXiv:2608.20905v1 Announce Type: new Abstract: Face-to-face audiovisual interaction is central to human communication, conveying rich emotional and social cues. However, existing multimodal dialogue...

By Yi Zheng, Yifan Xu, Yan Zhou, Hejia Chen, Chunyu Qiang, Xiaoqiang Liu, Xiaohan Li, Shenze Huang, Yue Zhang, Guoying Zhao, Pengfei Wan
multimodalsafety
More like this →
arXiv AI
Jul 2

SocialOmni: Benchmarking Audio-Visual Social Interactivity in Omni Models

arXiv:2603. 16859v2 Announce Type: replace Abstract: Omni-modal large language models (OLMs) redefine human-machine interaction by natively integrating audio, vision, and text.

By Tianyu Xie, Jinfa Huang, Yuexiao Ma, Rongfang Luo, Yan Yang, Wang Chen, Yuhui Zeng, Yixuan Zou, Qingchuan Ma, Zhiqiang Lu, Ruize Fang, Xiawu Zheng, Jiebo Luo, Rongrong Ji
llmsmultimodalbenchmarks
More like this →
arXiv AI
3d ago

Talking to Me or Someone Else? Rethinking Talk-to-Me Detection in Egocentric Videos

arXiv:2609.14118v1 Announce Type: cross Abstract: Online understanding of who is talking to the camera wearer is a key capability for egocentric social interaction. However, existing talk-to-me (TTM)...

By Feiyu Du, Xi He, Jia Li, Yapeng Tian, Weili Wu
multimodalbenchmarks
More like this →
arXiv AI
Jun 16

NVMOS: Non-Verbal Vocalization Quality Assessment in Speech

arXiv:2606. 15888v1 Announce Type: cross Abstract: Non-verbal vocalizations (NVs), such as laughter, sighs, and coughs, are important acoustic cues for emotion and intent.

By Jialong Mai, Jinxin Ji, Xiaofen Xing, Wencui Liu, Xiangmin Xu
llmsmultimodal
More like this →
About Pricing API Newsletter Sources Privacy Terms Refunds Accessibility Provider info Contact RSS

The Flow links to publishers and never republishes their articles. Summaries are machine-generated.

v1.1.0 · 5f852ea