Hugging Face Trending Papers

Guide Me Out: A Framework to Benchmark VLM Operators Communication in Crisis Scenarios

Effective crisis response requires spatially grounded communication that bridges linguistic guidance of civilians with the physical environment, accounting for structural bottlenecks, evolving threats, and agent-specific contexts. Yet, current NLP research in crisis communication remains mainly limited to static, text-only classification settings, overlooking the critical communicative role of AI operators in dynamic, embodied scenarios.

arXiv AI
Aug 18

Evaluating Multimodal LLMs across Text and Audio Modalities for Accessible Disaster Assistance

arXiv:2608. 14651v1 Announce Type: new Abstract: Effective disaster risk communication is a foundational humanitarian challenge, yet current emergency infrastructure fails to meet the needs of individuals with access and functional needs, including hard-of-hearing individuals, pregnant women, mothers with toddlers, and elderly individuals with dementia.

By Anuridhi Gupta, Samara Mansoor, Hemant Purohit
arXiv Machine Learning
Jun 17

Tacit Coordination of Large Language Models

arXiv:2601. 22184v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed in multi-agent settings that require coordination without communication, from human-AI interaction to safety-critical scenarios.

By Ido Aharon, Emanuele La Malfa, Michael Wooldridge, Sarit Kraus
arXiv AI
Jun 6

DisasterBench: A Multimodal Benchmark for UAV-Based Disaster Response in Complex Environments

arXiv:2606. 06217v1 Announce Type: cross Abstract: When a disaster unfolds, responders must answer not only what is happening, but also why it is happening, what will happen next, and what to do now, often from noisy low-altitude UAV views and under tight on-site compute constraints.

By Tan Zhang, Quanyou Li, Lu Zhang, Jun Liu, Xiaofeng Zhu, Ping Hu
arXiv Computer Vision
Sep 21

DisasterInsight: A Building-Centric Benchmark for Evaluating Vision--Language Models in Disaster Response

DisasterInsight is a building‑centric benchmark designed to evaluate vision‑language models (VLMs) for disaster response. Built on the xBD satellite dataset, it adds OpenStreetMap‑derived functional labels to 134,108 building instances and offers 15 task types, including instance assessment, scene counting, multi‑instance reasoning, and structured report generation. Experiments show that VLMs excel at visible damage detection but struggle with building function, multi‑instance reasoning, counting, and grounded reporting, and instruction tuning only partially mitigates these gaps.

By Sara Tehrani, Yonghao Xu, Leif Haglund, Amanda Berg, Gulnaz Zhambulova, Michael Felsberg
arXiv AI
Sep 15

Multilingual Agent System for Inclusive Wildfire Evacuation Guidance

The paper introduces BEACON, a multilingual agent system designed to deliver personalized wildfire evacuation guidance to users, including navigation routes, checklists, and a chatbot that adapts to the user’s language. It integrates real‑time fire perimeter, evacuation orders, shelter data, GPS, and NOAA weather to predict fire danger using an XGBoost model and triggers alerts with polygon‑avoidant routing when danger is likely. The interface automatically switches languages based on the user’s recent settings or chatbot interactions, ensuring inclusive communication for non‑English speakers.

By Shruti Kulkarni, Lynn Tong, Aditi Namboodiripad, Chelyah Miller, Helen Lin, Peeyush Patel, Bogdan Bistriceanu, Diane Myung-kyung Woodbridge