arXiv Computer Vision

Towards Practical Compression of 3D Gaussian Splatting

The paper introduces COSA-GS, a new compression method for 3D Gaussian Splatting that avoids spatial aggregation by using anchor-wise causal factorization. It builds a compact learnable anchor latent from geometry context and fuses it with the geometry context to create an anchor context for attribute coding, employing only linear transformations and activations. The method is trained with rate–distortion optimization, adaptive Gaussian pruning, and quantization-aware training to ensure bit‑exact entropy decoding across platforms, achieving state‑of‑the‑art compression performance with fast, consistent cross‑platform decoding.

arXiv Computer Vision
Aug 28

KISS-GS: 3D Gaussian Splatting Compression Kept Simple

KISS-GS is a modular compression pipeline for 3D Gaussian Splatting (3DGS) scenes that separates compression from training. It first compacts a vanilla 3DGS scene by 15.7× using state‑of‑the‑art pruning, then encodes the result into the SOG‑XT image‑based format, achieving an additional 6.6× reduction. Optional encoding‑aware fine‑tuning can further cut the size by 2.2×, yielding total reductions of 85× to 319× on standard benchmarks while enabling web‑native decoding.

By Wieland Morgenstern, Friedrich Elias Branschke, Florian Fleischmann, Adrian Szatmari, Paul Schlack, Florian Barthel, Peter Eisert, Anna Hilsmann
Hugging Face Trending Papers
Aug 3

StreamSplat: Streaming Feed-Forward 3D Gaussian Splatting

Feed-forward 3D Gaussian Splatting enables efficient novel-view synthesis without per-scene optimization, but most existing methods assume a fixed set of context views and process them jointly. This limits their applicability to online scenarios where calibrated views arrive sequentially and the scene must be updated causally.

arXiv Computer Vision
Aug 31

Non-Uniform Quantisation for 3DGS Compression

The paper introduces a non‑uniform quantisation scheme designed for 3D Gaussian Splatting (3DGS) models, addressing their high bitrate demands. By applying importance‑weighted quantisation and merging, the method adapts to the data distribution and removes post‑voxelisation redundancy. Experiments on benchmark datasets show state‑of‑the‑art compression performance, and the scheme is compatible with any point‑cloud representation, positioning it as a candidate for future MPEG 3DGS standardisation.

By Bert Van hauwermeiren, Patrice Rondao Alface, Adrian Munteanu
Hugging Face Trending Papers
Jul 7

GaussFusion: Towards Multimodal 3D Gaussian Pretraining

3D Gaussian Splatting provides an explicit representation that jointly models geometry and appearance, serving as a scalable foundation for 3D representation learning. Existing pre-training methods for Gaussian representations, such as masked Gaussian reconstruction, primarily capture local structures but offer limited semantic supervision.

Hugging Face Trending Papers
Jul 27

GenSplatCodec: Feed-Forward Gaussian Splatting Compression via One-Step Diffusion

Feed-forward 3D Gaussian Splatting (3DGS) enables scalable scene reconstruction without per-scene optimization, yet produces dense Gaussians that are costly to store and transmit. Existing feed-forward Gaussian compression methods formulate decoding as deterministic representation recovery, which becomes inadequate at low bitrates when high-frequency textures and view-dependent appearance are discarded.

arXiv Computer Vision
Sep 24

Visibility-Guided Structured Measure Flow for Class-Conditioned 3D Gaussian Generation

The paper introduces VISTA-GS, a visibility‑guided structured measure flow framework for generating class‑conditioned 3D Gaussian Splatting (3DGS) objects. It treats a 3DGS object as a structured Gaussian measure weighted by opacity, anisotropic covariance, and multi‑view visibility, and employs a visibility‑aware VAE to learn permutation‑invariant, variable‑size, rendering‑aware latent representations. The method includes a renderer‑consistent measure flow and structure‑preserving patch transport, achieving 60–72% improvements over the strongest baseline on the VISTA‑Obj30 dataset in geometry, appearance, view‑consistency, and speed.

By Yizhao Wang