arXiv Computer Vision By Yesheng Zhang, Xiang Dai, Xu Zhao, Chongyang Zhang

NaCR: Visual Localization via NeRF-aided Camera Ray Regression

Read the original on arXiv Computer Vision →

NaCR: Visual Localization via NeRF-aided Camera Ray Regression proposes a unified framework that integrates Neural Radiance Fields (NeRF) with Camera Ray Regression (CRR) to improve visual localization accuracy. The method enhances the CRR baseline with three simple improvements, augments training data by synthesizing novel views from a pre‑trained NeRF, and employs a closed‑loop supervision pipeline that back‑propagates photometric rendering errors to refine predicted camera rays. A two‑stage training curriculum ensures stable convergence, and experiments on indoor and outdoor benchmarks show competitive accuracy with validated component efficacy.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Computer Vision.

Hugging Face Trending Papers
Jul 2

NeoMap: Training-free Novel-View Synthesis from Single Images and Videos

We study the challenging problem of novel view video synthesis from single images or monocular videos. Existing methods, which operate under the assumption that pre-trained video models lack native novel view synthesis capability and enforce view alignment via camera conditioning, task-specific fine-tuning, or stepwise hard denoising guidance, often suffer from artifacts and compromised global scene consistency.

arXiv Computer Vision
Sep 7

ARC-Loc: Leveraging Azimuthal Ray Convergence as a Geometric Cue for Direct Cross-View Localization

ARC‑Loc introduces a new cross‑view localization method that bypasses heavy Bird’s‑Eye‑View transformations and external depth models. By converting ground keypoints into azimuthal rays on a satellite map and exploiting their convergence at the user’s location, the approach uses a minimal Azimuthal Ray Convergence solver and an ARC loss to directly match ground and satellite images. Experiments on VIGOR and KITTI show that ARC‑Loc achieves competitive accuracy while offering faster, memory‑efficient inference and easy integration with existing frameworks.

By Hyeongsik Kim, Mincheol Kim, Heejoon Moon, Je Hyeong Hong
arXiv Computer Vision
Sep 21

Refining Ground Truth Poses in Autonomous Driving Datasets via Neural Rendering

arXiv:2504.15776v2 Announce Type: replace Abstract: Public autonomous driving datasets underpin the training and benchmarking of perception, mapping, and localization algorithms, yet residual inaccur...

By Quentin Herau, Nathan Piasco, Moussab Bennehar, Luis Rold\~ao, Dzmitry Tsishkou, Bingbing Liu, Cyrille Migniot, Pascal Vasseur, C\'edric Demonceaux