← Back to all news
arXiv Computer Vision August 26, 2026 By Urwa Fatima, Mohammad Zohaib, Francesca Odone, Nicoletta Noceti

Human-Inspired Social Engagement Analysis via Interpretable Mutual Visual Attention

Read the original on arXiv Computer Vision →

The Flow has not summarised this story yet — read it at arXiv Computer Vision.

  • benchmarks

One email a morning, machine-written

One email a day, machine-written, one click to leave. We never share your address.

Related stories

arXiv AI
Jul 7

Attention Dynamics in Diffusion Models: A Visual Analytics Framework for Human-AI Collaboration

arXiv:2607. 02563v1 Announce Type: cross Abstract: Diffusion-based text-to-image models can synthesize complex and highly structured visual content, yet the emergence and evolution of semantic structure remain difficult to interpret.

By Yiran Xiao, George Legrady
diffusionbenchmarks
More like this →
arXiv AI
Aug 11

Walking through Discussions: A Mobile Visual Analytics System for In-Situ Group Discussion Analysis

arXiv:2608. 08617v1 Announce Type: new Abstract: Group discussion-based teaching is widely used to foster collaborative learning, yet teachers in physical classrooms often struggle to simultaneously monitor multiple groups and quickly diagnose a target group before intervening.

By Yiping Sun, Ziyao Kang, Wei Zeng, Minli Wu, Jiazhi Xia
More like this →
arXiv Computer Vision
4d ago

Socialality Anchors: Towards Group-bounded Trajectory Prediction

arXiv:2609.36852v1 Announce Type: new Abstract: Trajectory prediction is a key component for understanding human behavior patterns in dynamic scenes. Researchers have devoted substantial efforts to m...

By Ziqian Zou, Conghao Wong, Qinmu Peng, Xinge You
agentsbenchmarkssafety
More like this →
arXiv AI
Jun 29

Can LLMs Reason About Attention? Towards Zero-Shot Analysis of Multimodal Classroom Behavior

arXiv:2604. 03401v4 Announce Type: replace-cross Abstract: Understanding student engagement usually requires time-consuming manual observation or invasive recording that raises privacy concerns.

By Nolan Platt, Sehrish Nizamani, Alp Tural, Elif Tural, Saad Nizamani, Andrew Katz, Yoonje Lee, Nada Basit
llmsmultimodalsafety
More like this →
arXiv Machine Learning
Jun 3

Social Caption: Evaluating Social Understanding in Multimodal Models

arXiv:2601. 14569v2 Announce Type: replace-cross Abstract: Social understanding abilities are crucial for multimodal large language models (MLLMs) to interpret human social interactions.

By Leena Mathur, Bhaavanaa Thumu, Youssouf Kebe, Louis-Philippe Morency
llmsmultimodal
More like this →
arXiv AI
2d ago

Do Vision Language Models Understand Human Engagement in Games?

arXiv:2603.18480v2 Announce Type: replace-cross Abstract: Inferring human engagement from gameplay video is important for game design and player-experience research, yet it remains unclear whether vi...

By Ziyi Wang, Qizan Guo, Rishitosh Singh, Xiyang Hu
llmsragmultimodal
More like this →
About Pricing API Newsletter Sources Privacy Terms Refunds Accessibility Provider info Contact RSS

The Flow links to publishers and never republishes their articles. Summaries are machine-generated.

v1.1.0 · 5f852ea