arXiv Computer Vision

Calibrated Multichannel Monocular Ranging From Standardized License Plates With Metrology-Exact Validation

arXiv Machine Learning
Sep 14

A Multi-Vehicle Dataset with Camera, LiDAR, and Radar Sensors and Scanned 3D Models for Custom Auto-Annotation using RTK-GNSS

The paper introduces a multi-vehicle dataset that includes camera, LiDAR, and radar sensor data along with scanned 3D models of all vehicles. Each vehicle’s pose and continuous kinematics are provided via RTK‑GNSS, enabling precise knowledge of the dynamic surroundings at any time. The dataset supports single‑ and multi‑object recordings with seven target vehicles, allowing evaluation of measurement effects such as occlusion and reflections thanks to known vehicle surface normals.

By Philipp Berthold, Bianca Forkel, Mirko Maehlisch
arXiv Computer Vision
Sep 25

Smartphone-Based Method for Automated Speed Enforcement

The paper presents a smartphone-based system that uses computer vision to automatically estimate vehicle speed and identify vehicles by license plate, make/model, and color. Experiments on a Brazilian dataset and real-world recordings in Austin, Texas show moderate recognition rates: 46% for license plates, 60.8% for color, 48.6% for make, and 16.89% for make/model. The study also discusses legal, technological, and practical considerations for deploying such smartphone recordings in traffic enforcement.

By Keya Li, Jahnavi Malagavalli, Lamha Goel, Tong Wang, Kara M. Kockelman
arXiv Computer Vision
Sep 18

Open-vocabulary 3D object detection with promptable segmentation

The paper introduces an open‑vocabulary 3D object detection pipeline that uses a promptable segmentation model (SAM3) to generate instance masks from six surround‑view cameras. These masks are converted into metric 3D boxes, achieving up to 0.413 mAP/0.555 NDS without any training when supervised box geometry is borrowed at inference. The approach also improves a supervised LiDAR‑only detector by 0.034 mAP through a camera‑witness rule, demonstrating that measurement precision, not 2D detection, limits performance.

By \"Omer Faruk Deniz, Mustafa Taha Ko\c{c}yi\u{g}it
arXiv Computer Vision
Sep 24

SGDet3D++: Geometry-Grounded Semantics for 4D Radar and Camera 3D Object Detection

SGDet3D++ introduces a geometry‑grounded approach to 4D radar‑camera 3D object detection by explicitly conditioning evidence on evolving object hypotheses. It employs Anchor‑Grounded Semantic Retrieval, Geometry‑Consistent Anchor Refinement, and Doppler‑Verified Correspondence to filter and align semantic, geometric, and temporal cues before updating queries. The method achieves significant performance gains on OmniHD‑Scenes, ManTruckScenes, and TJ4DRadSet, with detailed ablations showing improvements in occlusion handling, target‑return purity, and motion consistency.

By Xiaokai Bai, Zhenyu Fan, Lianqing Zheng, Songkai Wang, Si-Yuan Cao, Hui-liang Shen
Hugging Face Trending Papers
Jun 22

Scene-agnostic ALS boresight self-calibration

ALS boresight calibration has relied for two decades on dedicated flight patterns over structured scenes containing planar surfaces of varied aspect and slope. While reliable, this approach imposes constraints on the scene content and operations, which limits its applicability to boresight recovery within routine mapping missions.

arXiv Computer Vision
Aug 24

Multi-Modal Traffic Sign Detection with Semantic Attributes for Autonomous Driving

The paper introduces a multi‑modal traffic sign detection framework that fuses camera and LiDAR data using an Intensity‑Aware Deformable Fusion module to align retro‑reflective LiDAR cues with visual features. It also presents a dual motion‑model tracker to handle non‑linear perspective changes and a semantic attribute classification pipeline that estimates occlusion, readability, sign embeddedness, and road relevance. Evaluated on a dataset covering more than 60 countries and 2,500 hours of driving, the system achieves an Object Miss Ratio of 0.49% across 221,068 sequences, indicating strong global generalization for autonomous driving.

By Meda Lazar, Sourab Sridhar, Shashwata Gupta, Alexandra Tripcea, Varun Ravi, Senthil Yogamani