arXiv:2606. 16271v1 Announce Type: cross Abstract: Unsupervised 3D seismic horizon tracking faces a key limitation: signal-based propagators provide accurate trace-level alignment but often fail near faults, whereas texture-driven deep models are more robust to discontinuities, typically at the cost of labeled data requirements and reduced trace-level precision.
By Alexandre Thouvenot, Lionel Boillot, Vincent Gripon
arXiv:2608.21136v1 Announce Type: new
Abstract: Recently, open-vocabulary zero-shot 3D scene understanding using vision foundation models has emerged as a promising alternative to data-intensive supe...
By Jie Xu, Na Zhao
The paper presents a framework that adapts general-purpose Vision Foundation Models (VFMs) to seismic data denoising using Parameter‑Efficient Fine‑Tuning with Low‑Rank Adaptation (LoRA). It introduces a kurtosis‑guided unsupervised test‑time adaptation module that updates only LoRA parameters to self‑calibrate for site‑specific noise without ground truth. Experiments on exploration seismic images and DAS data demonstrate that the approach matches or surpasses domain‑specific models and generalizes well to unseen cross‑site data.
By Jiahua Zhao, Umair bin Waheed, Jing Sun, Yang Cui, Nikos Savva, Eric Verschuur
arXiv:2609.10322v1 Announce Type: new
Abstract: Transferring the rich priors of large 2D foundation models to sparse 3D LiDAR remains challenging, as training native 3D foundation models at comparabl...
By Samed Do\u{g}an, Nico Leuze, Alfred Sch\"ottl
arXiv:2608.29609v1 Announce Type: new
Abstract: Semantic segmentation is a crucial task for understanding Mars, the most Earth-like planet in our solar system. However, it is challenging because the...
By Ming-Han Lee, Chi-Yeh Chen
arXiv:2609.36193v1 Announce Type: new
Abstract: Learning from scientific measurements often requires aligning modalities with different spatial support and resolution. Subsurface characterization is...
By Meher Gajula, Keyla Gonzalez, Ben Lasscock, Alejandro Valenciano
arXiv:2609.01172v1 Announce Type: new
Abstract: Monocular depth estimation has long stood as a fundamental challenge in computer vision, enabling a wide range of applications including 3D reconstruct...
By Muxin Liu, Xiaoyang Lyu, Yang-Tian Sun, Yi-Hua Huang, Ziyi Yang, Peng Dai, Xiaojuan Qi
arXiv:2606. 15786v1 Announce Type: cross Abstract: The advent of large pretrained foundation models for computer vision has significantly improved the efficiency of visual data interpretation.
By Aniq Ahmad, Heather Bedle, Ahmad Mustafa
MAETrack introduces a lightweight framework to adapt pretrained masked autoencoder (MAE) representations for 3D single object tracking (SOT). It uses Layer‑Selective Initialization (LSI) to keep shallow geometric layers from the pre‑training while re‑initializing deeper layers, and Geometric Residual Gating (GRG) to emphasize salient regions in BEV features before template‑search fusion. Experiments on standard 3D SOT benchmarks show consistent improvements over vanilla fine‑tuning with minimal computational cost.
By Sifan Zhou, Qiwei Wang, Linyue Tan, Ziyu Liu, Ziyu Zhao, Xiaobo Lu
arXiv:2607. 12433v1 Announce Type: cross Abstract: Diffusion models have recently become the dominant paradigm for monocular depth estimation (MDE).
By Zijie Wang, Wei Zhang, Weiming Zhang, Xiao Tan, Weikai Chen, Xiaoxu Li, Guanbin Li
arXiv:2609.22896v1 Announce Type: new
Abstract: Autonomous vehicles operating in open-world scenarios are inevitably confronted with previously unknown objects, such as exotic animals or loose cargo....
By Serin Varghese, Fabian H\"uger, Kira Maag
arXiv:2609.13332v1 Announce Type: new
Abstract: Automated landslide segmentation on Mars is one of the important tasks for understanding its surface processes, and all will aid in future space explor...
By Leo Thomas Ramos, Sidike Paheding, Abel A. Reyes-Angulo, Rajaneesh A., Sajinkumar K. S., Angel D. Sappa, Thomas Oommen