arXiv Computer Vision

CoLMIN: LLM-based Multi-Decision Path Negotiation for Cooperative Autonomous Driving

CoLMIN is an LLM-based framework for cooperative autonomous driving that addresses premature convergence to suboptimal solutions in multi-solution traffic scenarios. It introduces a Multi-Intent Negotiation module that generates multiple candidate driving intentions, an Evaluation-based Shallow Reflection Module that provides feedback to accelerate consensus, and a Deep Reflection Module that mitigates cognitive fixation by reflecting on negotiation histories. Experiments in the CARLA simulation show that CoLMIN outperforms existing methods in challenging interactive driving scenarios.

Hugging Face Trending Papers
Sep 4

CoLMIN: LLM-based Multi-Decision Path Negotiation for Cooperative Autonomous Driving

CoLMIN is an LLM-based framework for multi-decision path negotiation in cooperative autonomous driving. It introduces a Multi-Intent Negotiation module that generates multiple driving intentions, an Evaluation-based Shallow Reflection Module that provides feedback to accelerate consensus, and a Deep Reflection Module that mitigates cognitive fixation by reflecting on negotiation histories. Experiments in the CARLA simulation show that CoLMIN outperforms existing methods in challenging interactive driving scenarios.

arXiv AI
Aug 11

CMU-Drive and V2V-VLA: Cooperative Multi-agent Unified Driving with Reasoning Benchmark and Vehicle-to-Vehicle Vision-Language-Action Models

arXiv:2608. 07621v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have recently achieved impressive performance for end-to-end autonomous driving, yet existing approaches are primarily designed for an individual single autonomous driving agent with limited support for cooperative perception, reasoning, and planning.

By Hsu-kuang Chiu, Stephen F. Smith
arXiv Machine Learning
Jun 4

CADET: A Modular Platform for Evaluating Distributed Cooperative Autonomy in Connected Autonomous Vehicles

arXiv:2606. 04072v1 Announce Type: cross Abstract: Deep learning models are increasingly central to autonomous vehicle (AV) pipelines, yet their integration has traditionally followed a monolithic design where perception, planning, and control execute on a single onboard computer.

By Pragya Sharma, Brian Wang, Mani Srivastava
arXiv Computer Vision
Sep 3

VIPS: Vehicle-Infrastructure Cooperative Planning Benchmark via Pseudo-Simulation

VIPS is a benchmark for vehicle‑to‑infrastructure cooperative autonomous driving that uses pseudo‑simulation to combine vehicle and infrastructure observations, enabling scalable yet realistic evaluation of robustness and error propagation without full simulation. The paper also introduces CoS‑V2X, a cooperative planning framework that employs sparse representations to model vehicle‑infrastructure interactions efficiently and robustly under heterogeneous observations.

By Hoonhee Cho, Jae-Young Kang, Giwon Lee, Hyemin Yang, Heejun Park, Kuk-Jin Yoon
arXiv AI
Jul 24

Compact Latent Coordination for Autonomous Vehicles at Unsignalized Intersections

arXiv:2607. 21488v1 Announce Type: cross Abstract: Coordinating autonomous vehicles at unsignalized intersections remains a critical challenge for multi-agent reinforcement learning (MARL) systems, which typically struggle with combinatorial action spaces, reliance on privileged information, or rigid agent designs.

By Gil Lifshits, Igal Bilik, Gilad Katz