arXiv AI By Xiangwen Zhang, Qian Zhang, Longfei Han, Qiang Qu, Xiaoming Chen, Weidong Cai

AccidentSim: Generating Vehicle Collision Videos with Physically Realistic Collision Trajectories from Real-World Accident Reports

Read the original on arXiv AI →

AccidentSim is a framework that generates physically realistic vehicle collision videos by extracting physical clues from real-world accident reports. It uses a reliable physical simulator to replicate post-collision trajectories, builds a trajectory dataset, fine‑tunes a language model to predict consistent trajectories from user prompts, and finally renders high‑quality videos with Neural Radiance Fields. The resulting videos show strong visual and physical authenticity compared to existing methods.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv AI
Jul 10

AUTOPILOT VQA: Benchmarking Vision-Language Models for Incident-Centric Dashcam Understanding

arXiv:2607. 08745v1 Announce Type: new Abstract: Recent advances in Vision-Language Models, Large Language Models, and Multimodal Large Language Models have improved autonomous driving tasks such as scene understanding, decision making, trajectory prediction, and visual question answering.

By Siddharth Damodharan, Radhika Gupta, Ali Alshami, Ryan Rabinowitz, Jugal Kalita
arXiv AI
Sep 18

LLM-Guided Transformation of Non-Critical Driving Scenes into Safety-Critical Scenarios Using Augmented Reality

The paper introduces an automated pipeline that converts non‑critical driving scenes into safety‑critical scenarios by integrating computer vision, Large Language Models (LLMs), and Augmented Reality (AR). It detects and tracks road users, extracts safety features such as distance, velocity, motion direction, and Time‑to‑Collision (TTC), and evaluates scene criticality. Safe scenes are then modified by an LLM, which generates realistic collision‑inducing objects and behaviors that are overlaid onto the original scene using AR, achieving 97.52% safety classification accuracy on the nuScenes dataset and producing realistic scenarios like pedestrian crossings, rear overtaking vehicles, and sudden‑stop events.

By Noura Fady, Farah Khaled, Catherine M. Elias
arXiv AI
Sep 3

CrashDiffuser: VLM-Guided Collision Intent Reasoning for Fine-Grained Safety-Critical Traffic Scenario Generation

CrashDiffuser is a closed-loop VLM‑guided diffusion framework designed for fine‑grained safety‑critical traffic scenario generation. It separates semantic collision reasoning from trajectory synthesis via a hierarchical collision‑intent interface that specifies target contact regions (head, rear, or side). The system uses a vision‑language model to extract scene context and predict structured action tuples, which condition a diffusion model to produce executable adversarial trajectories, achieving high target‑collision and contact‑region control rates on WOMD‑derived scenarios.

By Shucheng Zhang, Yuang Zhang, Bingzhang Wang, Muhammad Monjurul Karim, Kehua Chen, Yinhai Wang