arXiv Computer Vision By Han-Gyeol Kim, JaeWan Park, Junmin Park, Darongsae Kwon

SatUnreal: A High-Precision Synthetic Dataset for Satellite Stereo Matching via Unreal Engine

Read the original on arXiv Computer Vision →

SatUnreal is a synthetic dataset created with Unreal Engine that offers 10,000 high‑resolution (0.3 m GSD) satellite stereo pairs. It addresses key limitations of existing benchmarks by ensuring physical geometry simulation, spatio‑temporal consistency, topographic diversity, and mathematically precise occlusion masks via a two‑step line‑trace algorithm. Models trained solely on SatUnreal outperform those trained on real datasets when transferred to real‑world benchmarks such as US3D and WHU‑Stereo.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Computer Vision.

Hugging Face Trending Papers
Jun 25

SatSplatDiff: Geometry-preserving generative refinement for high-fidelity satellite Gaussian Splatting

Gaussian Splatting has been recently explored for satellite 3D reconstruction, demonstrating flexibility and efficiency in representing radiometrically diverse satellite scenes. However, the limited top viewpoint of satellite imagery results in insufficient supervision on building facades, leaving surface holes and degraded visual fidelity.

arXiv Computer Vision
Aug 28

SpatialCrafter: Single Image World Modeling with Generative 3D Proxies

SpatialCrafter introduces a two‑stage framework for single‑image world modeling that first generates a global 3D proxy using a Point‑anchored Sparse Structure Flow module, then refines appearance with a Generative Deferred Refiner built on a video diffusion model. The method incorporates Parallel Geometry Injection and Proxy‑Aware Corruption training to integrate the proxy without disrupting the pretrained generative manifold, and it is evaluated on a newly constructed dataset of 115K scenes. Experiments demonstrate that SpatialCrafter outperforms existing approaches, reducing long‑term drift and maintaining consistency under rapid camera motion and extreme viewpoints.

By Chuan Fang, Lingteng Qiu, Yixun Liang, Rui Chen, Kunming Luo, Zhaohua Zheng, Tongyuan Bai, Feipeng Tian, Zilong Dong, Zihan Zhou, Ping Tan