arXiv Computer Vision

LiTe-GS: Oracle-Efficient Next Best View Selection for 3D Gaussian Splatting

arXiv:2609. 30393v1 Announce Type: new Abstract: Selecting informative camera views is critical for efficient training and adaptive refinement in 3D Gaussian Splatting, where each observation significantly influences model parameters.

arXiv Computer Vision
4d ago

Gauss What You Need: Compact Gaussian Splatting Across Scene Scales

Gauss What You Need: Compact Gaussian Splatting Across Scene Scales introduces TangoGS, a method that automatically selects the number of Gaussian primitives for 3D Gaussian Splatting by combining capture-derived model sizing with training-based adaptation. The approach first estimates a learning allowance based on the capture’s total pixels, then adjusts the number of Gaussians during training according to reconstruction quality. On standard benchmarks, TangoGS matches the best baseline’s PSNR while using 48% fewer Gaussians, and on larger captures it scales automatically to achieve the highest mean PSNR with 2.3× more Gaussians.

By Afif Boudaoud, Jiayi Liu, Alexandru Calotoiu, Torsten Hoefler
arXiv Computer Vision
1d ago

Matisse: Evidence-Space Reasoning for Active 3D Reconstruction

Matisse is a training‑free framework that combines active 3D reconstruction with keyframe selection by using evidence from a pretrained generative 3D model. It estimates evidential uncertainty via cross‑attention on 3D latent tokens and derives an evidential information gain to guide view acquisition and keyframe selection, reducing redundant observations and supporting multi‑object scenes with occlusion‑aware aggregation. On GSO30, YCB‑V, and Replica, Matisse improves Chamfer distance by 12.7%, 3.8%, and 9.2% respectively, and speeds up end‑to‑end reconstruction by 1.5× compared to the best baseline.

By Xihang Yu, Kaichen Zhou, Lorenzo Shaikewitz, Cl\'ement Jambon, Xiao Zhan, Rajat Talak, Luca Carlone
arXiv Computer Vision
Sep 21

2D GauSS-MI: Efficient Active Scene Reconstruction with Balanced Visual and Geometric Quality

The paper introduces 2D GauSS-MI, an active scene reconstruction framework that uses 2D Gaussian Splatting (2DGS) to efficiently process incremental RGB‑D data. It presents an online 2DGS mapping pipeline and a probabilistic reliability model to assess view‑dependent reconstruction quality. Leveraging this model, the authors define a Shannon Mutual Information metric that guides active view selection, balancing visual and geometric quality while reducing computational cost and storage compared to state‑of‑the‑art baselines.

By Yuhan Xie, Jia Pan
arXiv AI
Aug 20

GS-VLA: Plug-and-Play Viewpoint Canonicalization for Frozen VLA Policies via Gaussian Splatting

GS‑VLA introduces a lightweight, plug‑and‑play framework that uses a 4 M‑parameter 3D‑Gaussian canonicalizer to adapt frozen Vision‑Language‑Action (VLA) policies to viewpoint shifts without retraining the policy. By treating viewpoint changes as a localized novel‑view synthesis problem under a locality assumption, the method normalizes observations through a scene‑ and policy‑independent disocclusion task. Experiments on the LIBERO benchmark demonstrate that GS‑VLA recovers a large portion of performance lost due to camera displacement, improving results across different policy architectures, unseen task suites, and perturbation scales. whyItMatters":"The approach offers a computationally efficient alternative to costly fine‑tuning or generative augmentation, enabling robust VLA deployment in real‑world settings where camera configurations may vary."

By Yechan Park, HyunJin Kim