arXiv AI

DA-NBV: A Direction-Aware Next-Best-View Planner for Efficient 3D Reconstruction of Ships at Sea

arXiv:2608. 08025v1 Announce Type: cross Abstract: Accurate 3D reconstruction of ships at sea is important for maritime supervision, damage assessment, and autonomous maritime operations.

arXiv Computer Vision
Sep 11

Diagnosing and Dynamically Filtering Occupancy World Models for Active Mapping

The paper investigates how inaccuracies in pretrained occupancy networks affect active mapping robots that select camera viewpoints to reconstruct unknown 3D scenes. By fixing the planner and varying the occupancy representation—ranging from no completion to ground‑truth occupancy—the authors find that correcting false positives or false negatives alone does not reliably improve coverage, highlighting a disconnect between occupancy accuracy and planning performance. They propose a dynamic filtering strategy that retains predictions in unexplored space while suppressing unsupported occupancy based on online observations, which preliminarily shows it can steer viewpoint selection toward reachable surfaces that would otherwise remain unseen.

By Jiahui Zhang, Gongbo Liang, Yu Zhang
arXiv Machine Learning
Aug 3

ASVSim (AirSim for Surface Vehicles): A High-Fidelity Simulation Framework for Autonomous Surface Vehicle Research

arXiv:2506. 22174v3 Announce Type: replace-cross Abstract: The transport industry has recently shown significant interest in unmanned surface vehicles (USVs), specifically for port and inland waterway transport.

By Bavo Lesy, Siemen Herremans, Robin Kerstens, Jan Steckel, Walter Daems, Siegfried Mercelis, Ali Anwar
arXiv Computer Vision
3d ago

Matisse: Evidence-Space Reasoning for Active 3D Reconstruction

Matisse is a training‑free framework that combines active 3D reconstruction with keyframe selection by using evidence from a pretrained generative 3D model. It estimates evidential uncertainty via cross‑attention on 3D latent tokens and derives an evidential information gain to guide view acquisition and keyframe selection, reducing redundant observations and supporting multi‑object scenes with occlusion‑aware aggregation. On GSO30, YCB‑V, and Replica, Matisse improves Chamfer distance by 12.7%, 3.8%, and 9.2% respectively, and speeds up end‑to‑end reconstruction by 1.5× compared to the best baseline.

By Xihang Yu, Kaichen Zhou, Lorenzo Shaikewitz, Cl\'ement Jambon, Xiao Zhan, Rajat Talak, Luca Carlone
arXiv Computer Vision
Sep 25

OceanXL: Large-scale Underwater 3D Gaussian Splatting via Block Partitioning and Adaptive Pruning

OceanXL is a new framework that applies 3D Gaussian Splatting to large-scale underwater scenes by partitioning them into spatially coherent blocks and using adaptive pruning to remove redundant primitives. This divide‑and‑conquer approach improves training efficiency and rendering performance while maintaining global geometric consistency. The authors also release a large underwater dataset and demonstrate that OceanXL achieves favorable scalability, compactness, and efficiency compared to existing baselines, with competitive quality on smaller datasets and smaller model sizes than other underwater methods.

By Haoran Wang, Shaoyu Cai, Adrian Azzarelli, Zhuodong Jiang, Guoxi Huang, Eng Tat Khoo, Brett Seymour, Fan Zhang, David Bull, Nantheera Anantrasirichai
arXiv Computer Vision
Sep 7

AquaBEV: Monocular Underwater BEV Occupancy with 3D Sonar Supervision

AquaBEV is a monocular underwater occupancy model that predicts local bird’s‑eye‑view (BEV) occupancy from a single RGB image. It uses paired 3D imaging sonar data as geometric supervision during training, mapping visual features into a calibration‑free polar representation and decoding along the range dimension before reconstructing Cartesian BEV coordinates. In a controlled underwater occupancy benchmark, AquaBEV outperforms the strongest transferred baseline with 31.4 % Visible IoU and 38.6 % Observed IoU, achieving 4.0 % and 4.3 % relative improvements respectively.

By Trung Tien Dong, Shengji Jin, Chen Chen, Yi Sheng, Xiaomin Lin