arXiv:2607.16862v2 Announce Type: replace
Abstract: LiDAR place recognition supports loop closure, relocalization, and multi-agent map management. As robotic platforms increasingly combine LiDARs wit...
By Nikolaos Stathoulopoulos, George Nikolakopoulos
arXiv:2608. 19522v1 Announce Type: cross Abstract: Scan-to-map LiDAR odometry drifts unboundedly along the unobservable axes of geometrically degenerate environments like tunnels and corridors, and existing degeneracy handling requires environment-specific parameter tuning.
By Eunsoo Im
3D object detection is the backbone of perception for automated vehicles (AV) and broader intelligent transportation systems applications. Long-range detection is challenging because sensing evidence is sparse; yet this ``long-range'' scenario is routine in traffic.
arXiv:2606. 09634v1 Announce Type: cross Abstract: 3D object detection is the backbone of perception for automated vehicles (AV) and broader intelligent transportation systems applications.
By Debojyoti Biswas, Xianbiao Hu
arXiv:2607.06782v2 Announce Type: replace-cross
Abstract: Under field-of-view (FOV) mismatch, pooling LiDAR features over unequal angular support can distort compact retrieval keys and exclude correc...
By Jinseop Lee
SlugTrails is a new egocentric benchmark for floor‑plan‑based indoor visual localization in large buildings, featuring 30 Hz Aria glasses recordings across three campus buildings and six floors (22 089 m²). The dataset includes CAD‑derived floor plans with semantic classes, circulation masks, and laser‑surveyed anchors, and supports three realistic sensing protocols: single walking frames, stationary multi‑view sweeps, and walking streams with odometry. Evaluation of five geometric and learned systems shows that stock models perform poorly, but fine‑tuning on SlugTrails significantly improves performance and cross‑dataset generalization, indicating that data scarcity limits current localization methods.
arXiv:2607. 02561v1 Announce Type: cross Abstract: Consumer depth sensors such as the LiDAR scanner on recent iPhones provide metric range, but their useful range is short and their returns are sparse.
By Jinwen Wen
SlugTrails is a new egocentric benchmark for floor‑plan‑based indoor visual localization in large buildings, featuring 30 Hz Aria glasses recordings across three campus buildings and six floors. The dataset includes CAD‑derived floor plans with semantic classes, circulation masks, and laser‑surveyed anchors for trajectory alignment. Five representative systems were evaluated, showing that stock checkpoints perform poorly while fine‑tuning on SlugTrails significantly improves performance and cross‑dataset generalization, indicating that data scarcity limits current methods.
By Yunqian Cheng, Roberto Manduchi
DXPR is a depth‑based cross‑modal place recognition framework that matches monocular camera queries to a LiDAR map using a single vision foundation model backbone. By converting both modalities into a unified depth image representation, DXPR learns modality‑invariant global descriptors without modality‑specific encoders. A geometry‑aware overlap miner refines pairwise metric learning by computing pixel‑level overlap scores, and extensive tests on KITTI and Boreas show strong performance across seasons, weather, and day/night conditions, outperforming prior CMPR baselines.
arXiv:2609.38864v1 Announce Type: new
Abstract: Embodied tasks demand accurate, flexible, and semantically rich 3D scene representations. 3D semantic occupancy is well suited to this requirement, as...
By Jinglong Wang, Yunjie Wang, Zhiyang Zhang, Jiawei He, Ye Yuan, Bo Qiu, Jing Zhang
arXiv:2608. 07579v1 Announce Type: cross Abstract: The AI City Challenge 2026 Track 1 evaluates multi-camera 3D perception in large indoor warehouses under a synthetic-to-real (Sim2Real) setting; depth is available only for training and validation, so inference is RGB-only.
By Abdullah Naeem, Anav Katwal, Ayon Dey, Noman Khan, Md Tamjidul Hoque
Scene coordinate regression (SCR) achieves strong performance in outdoor LiDAR localization, but it usually requires scene-specific training that can take days, limiting practical deployment. Recent w...