arXiv Computer Vision

Frequency-Decomposed Avatar Representation for Varying Camera Distances

arXiv Computer Vision
6d ago

NBAvatar: Neural Billboards Avatars with Realistic Hand-Face Interaction

NBAvatar is a method for realistic rendering of head avatars that handles non‑rigid deformations caused by hand‑face interaction. It introduces a hybrid implicit‑explicit representation, combining explicit oriented planar primitives with implicit neural rendering, and uses a geometry‑aware training scheme to jointly optimize these representations. The approach achieves up to 53% LPIPS reduction compared to Gaussian‑based avatar methods, improves PSNR and SSIM, and surpasses the state‑of‑the‑art InteractAvatar in structural similarity for novel‑view and novel‑pose rendering.

By David Svitov, Mahtab Dahaghin, Pietro Morerio, Alessio Del Bue
Hugging Face Trending Papers
Jun 21

Generative Relightable Avatars

We present Generative Relightable Avatars (GRA), a person-specific method for photorealistic free-view rendering and environment-map relighting of full-body humans. We postulate that modeling fine-grained appearance details is inherently a one-to-many problem that can benefit from a generative formulation.

arXiv Computer Vision
Sep 18

AGORA: Adversarial Generation Of Real-time Animatable 3D Gaussian Head Avatars

AGORA is a new framework that extends 3D Gaussian Splatting with a generative adversarial network to produce high‑fidelity, animatable 3D head avatars. It introduces a lightweight FLAME‑conditioned deformation branch that predicts per‑Gaussian residuals for identity‑preserving, fine‑grained expression control, and a dual‑discriminator training scheme that enforces expression fidelity. The system achieves real‑time inference at 250 FPS on a single GPU and, for the first time, CPU‑only animatable 3DGS avatar synthesis at ~9 FPS.

By Ramazan Fazylov, Sergey Zagoruyko, Aleksandr Parkin, Stamatis Lefkimmiatis, Ivan Laptev
arXiv Computer Vision
Sep 11

Revisiting Avatar-As-Image: High-Fidelity Registration is All You Need

The paper introduces AvaImg, a multi‑stage optimization pipeline that achieves high‑fidelity SMPL(-X)+D registrations with UV texture for arbitrary clothed scans. By enforcing a body‑inside‑clothing constraint through signed winding numbers and employing a three‑level efficiency cascade, AvaImg significantly reduces runtime and storage while recovering fine surface detail via coarse‑to‑fine displacement optimization. The resulting textured registrations are nearly indistinguishable from scans, and encoding the UV maps with a frozen FLUX VAE demonstrates compatibility with 2D generative models, enabling 3D avatar generation using image‑based priors.

By Margaret Kostyrko, Yuxuan Xue, Garvita Tiwari, Gerard Pons-Moll
arXiv Computer Vision
4d ago

AESplat: Advancing Pose-Free Feed-Forward 3D Gaussian Splatting via Decoupled Appearance Modeling

AESplat is a new pose‑free feed‑forward 3D Gaussian Splatting framework that improves rendering quality by decoupling view‑independent and view‑dependent appearance modeling. It directly extracts the base view‑independent appearance from input images and predicts higher‑order spherical harmonic coefficients with a shallow MLP that incorporates 3D‑aware inductive biases. Experiments on several datasets show AESplat outperforms state‑of‑the‑art methods, achieving up to 0.8 dB higher PSNR than NAS3R and 1.1 dB over DepthSplat on RealEstate10K.

By Shiwei Ren, Zhiang Liu, Yongchun Fang, Hongwei Chen