arXiv Computer Vision By Daniele Lizzio Bosco, Shuteng Wang, Giuseppe Serra, Vladislav Golyanik

Quantum Implicit Neural Representations for Novel View Synthesis

Read the original on arXiv Computer Vision →

Quantum Implicit Neural Representations for Novel View Synthesis introduces 3D Quantum Implicit Scene Representation (3D-QISR), a hybrid quantum‑classical radiance‑field model that replaces the classical NeRF backbone with parameterised quantum circuits. Two architectures are proposed: Full 3D-QISR, which uses a unified quantum state, and Dual‑Branch 3D-QISR, which separates spatial and view‑dependent embeddings to reduce complexity and improve scalability. Experiments on moderate‑resolution novel‑view synthesis benchmarks show that 3D‑QISR on simulated quantum hardware matches or outperforms classical baselines while using fewer than half the trainable parameters.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Computer Vision.

arXiv Computer Vision
Sep 3

RoGe: Novel View Synthesis via End-to-End Implicit Reconstruction and Generation

RoGe is a new end‑to‑end framework for novel view synthesis that jointly learns an implicit 3D scene representation and a video diffusion model. It eliminates the need for explicit 3D intermediates by querying the implicit scene with camera rays to produce geometric features that condition the diffusion model. Experiments on DL3DV show that RoGe surpasses reconstruction‑based, generation‑based, and hybrid baselines in image quality and temporal consistency, and ablations confirm the benefits of ray‑queried features and joint training.

By Xiaolei Lang, Ze Kang, Zehao Huang, Naiyan Wang
arXiv AI
4d ago

Spackle: Completing Large View Single Image NVS with Adaptive Gaussians

Spackle is a lightweight residual learning framework designed to improve large-view single-image novel view synthesis (NVS) by mitigating capacity competition in hybrid decoupled systems that combine 3D Gaussian Splatting (3DGS) and diffusion models. It operates in three stages: predicting base 3DGS attributes, automatically identifying poorly reconstructed regions, and learning a residual 3DGS focused on those areas. During inference, Spackle merges the baseline and augmented Gaussians to produce high-fidelity novel views, achieving state‑of‑the‑art performance on large-view-deviation cases.

By Xuanzhi Liu, Yuhe Zhou, Xinyi Wu, Zhenyao Wu, Jinghao Chen, Ruize Han, Song Wang