arXiv Machine Learning

Beyond Direct Sensing: Harnessing Indirect Observations from Third-Party Sensors in Vehicle Tracking

arXiv Computer Vision
5d ago

MPT: Missing Prototype Tracking via Barycentric Reconstruction in Vehicular Federated Learning

arXiv:2609.12771v1 Announce Type: cross Abstract: Cross-vehicle federated learning enables vehicles to collaboratively improve perception models while keeping locally collected driving data private....

By Hanju Jang (Yonsei University), Gyeongmin Han (Yonsei University), Sungmin Lee (Yonsei University), Kichang Lee (Yonsei University), Chunghan Lee (Toyota Motor Corporation), JeongGil Ko (Yonsei University)
arXiv Computer Vision
Aug 31

When the City Teaches the Car: Label-Free 3D Perception from Infrastructure

The paper proposes a label‑free 3D perception framework where roadside units (RSUs) act as unsupervised teachers for self‑driving cars. RSUs learn local 3D detectors from unlabeled data and broadcast predictions to passing vehicles, which use these as pseudo‑labels to train an ego‑centric detector. In a CARLA simulation, the method achieves 82.3% AP for vehicle detection, approaching a fully supervised upper bound of 94.4%, and demonstrates scalability and complementarity with existing ego‑centric approaches.

By Zhen Xu, Jinsu Yoo, Cristian Bautista, Zanming Huang, Tai-Yu Pan, Zhenzhen Liu, Katie Z Luo, Mark Campbell, Bharath Hariharan, Wei-Lun Chao
Hugging Face Trending Papers
Aug 17

RadioVIL: Anomaly-Aware Diffusion Models for Radio Map Inpainting and Zero-Shot Vehicle Localization

High-precision radio map construction is essential for emerging 6G Integrated Sensing and Communication (ISAC) applications, including digital twins and intelligent transportation. However, existing deep learning methods predominantly treat this as a pure image completion task, resulting in over-smoothed reconstructions that fundamentally erase high-frequency scattering signatures of dynamic physical entities such as hidden vehicles.

arXiv Computer Vision
3d ago

GRACE: Geometry- and Ray-Aware Camera-Efficient Multi-View Pedestrian Tracking

GRACE is a camera‑efficient multi‑view pedestrian tracker that reduces the number of required cameras while maintaining high tracking accuracy. It combines volumetric‑guided fusion of homography‑based BEV features with 3D‑lifted features, uses ray conditioning to incorporate each camera’s viewing direction, and employs BEV Track Recovery to continue existing tracks with low‑confidence detections. On the WildTrack dataset, GRACE raises MOTA from 83.54 to 91.07 compared to the baseline TrackTacular.

By Taigo Sakai, Kazuhiro Hotta, Hiroki Kouno, Naoki Kato
arXiv AI
Aug 11

ATLASFusion: Aggregation Tracking with Location-Aware Sparse Fusion for Robust Spatio-Temporal Multi-View Pedestrian Tracking

arXiv:2509. 08421v2 Announce Type: replace-cross Abstract: For multimedia spatial intelligence through time, multi-view multi-object tracking (MVMOT) suffers from persistent challenges in maintaining consistent object identities across different camera views, leading to tracking inaccuracies.

By Keisuke Toida, Taigo Sakai, Takeshi Nakamura, Hiroshi Shimizu, Kazuhiro Hotta
Hugging Face Trending Papers
Jul 15

PiVoT: A Variational Solution for Real-time Large-scale Multi-object Detection and Tracking under Heavy Clutter

Multi-object detection and tracking from noisy point clouds remain challenging in many data-scarce radar applications. Current Bayesian trackers based on Poisson measurement models offer a training-free solution but struggle to achieve accuracy and efficiency under severe clutter, large object populations, and full-resolution Doppler point clouds.

arXiv Machine Learning
1d ago

QoS-Aware Federated Learning for Multimodal In-Cabin Interaction in Smart Vehicles

The paper introduces FedQoS, an asynchronous federated learning framework designed for multimodal in‑cabin interaction in smart vehicles. It uses a two‑phase gating mechanism: a resource‑aware training gate that starts local learning only when sensing buffers and energy reserves meet safety thresholds, and a QoS‑aware transmission policy that gates uplink updates based on an efficiency score balancing model novelty, latency, and energy costs. Experiments on vehicular datasets show FedQoS achieves competitive personalized accuracy with only marginal loss compared to FedAvg, while reducing communication overhead by 76.7% and latency cost by 26.0%.

By Baran Can G\"ul, Mert Nak{\i}p, Nasser Jazdi, Michael Weyrich