SatSplatDiff: Geometry-preserving generative refinement for high-fidelity satellite Gaussian Splatting
Read the original on arXiv Computer Vision →The Flow has not summarised this story yet — read it at arXiv Computer Vision.
The Flow has not summarised this story yet — read it at arXiv Computer Vision.
Gaussian Splatting has been recently explored for satellite 3D reconstruction, demonstrating flexibility and efficiency in representing radiometrically diverse satellite scenes. However, the limited top viewpoint of satellite imagery results in insufficient supervision on building facades, leaving surface holes and degraded visual fidelity.
arXiv:2606.24138v2 Announce Type: replace Abstract: Generating explicit textured 3D city assets from a single satellite image is important for urban simulation and digital twins. Most prior methods,...
SpatialCrafter introduces a two‑stage framework for single‑image world modeling that first generates a global 3D proxy using a Point‑anchored Sparse Structure Flow module, then refines appearance with a Generative Deferred Refiner built on a video diffusion model. The method incorporates Parallel Geometry Injection and Proxy‑Aware Corruption training to integrate the proxy without disrupting the pretrained generative manifold, and it is evaluated on a newly constructed dataset of 115K scenes. Experiments demonstrate that SpatialCrafter outperforms existing approaches, reducing long‑term drift and maintaining consistency under rapid camera motion and extreme viewpoints.
arXiv:2608.23549v1 Announce Type: new Abstract: Rendering views using 3D scene representations such as Gaussian Splatting (3DGS), Neural Radiance Fields (NeRF), meshes, or even point clouds produces...
SatUnreal is a synthetic dataset created with Unreal Engine that offers 10,000 high‑resolution (0.3 m GSD) satellite stereo pairs. It addresses key limitations of existing benchmarks by ensuring physical geometry simulation, spatio‑temporal consistency, topographic diversity, and mathematically precise occlusion masks via a two‑step line‑trace algorithm. Models trained solely on SatUnreal outperform those trained on real datasets when transferred to real‑world benchmarks such as US3D and WHU‑Stereo.
arXiv:2609.10531v1 Announce Type: new Abstract: Image-to-3D models can generate visually compelling 3D assets from a single RGB image, but their geometry is often only loosely constrained by the avai...