arXiv:2606. 30347v1 Announce Type: cross Abstract: We present FFAvatar, a Transformer-based 3D Gaussian framework for fast construction of high-quality and animatable 4D head avatars from one or more reference portrait images.
By Jianjiang Yao, Ke Xian, Renxiang Dai, Robert Caiming Qiu
arXiv:2512.03593v2 Announce Type: replace
Abstract: We present a CloseUpAvatar - a novel approach for articulated human avatar representation supporting a wider range of camera motions, while preserv...
By David Svitov, Pietro Morerio, Lourdes Agapito, Alessio Del Bue
arXiv:2609.12850v1 Announce Type: new
Abstract: Accurate head modeling requires a stable yet expressive geometric representation. Existing Gaussian-based head avatars commonly rely on parametric temp...
By Lei Shi, Sen Peng, Zhiyang Deng, Zhonggui Chen, Xiaohu Guo, Baorong Yang, Xiao Dong
AGORA is a new framework that extends 3D Gaussian Splatting with a generative adversarial network to produce high‑fidelity, animatable 3D head avatars. It introduces a lightweight FLAME‑conditioned deformation branch that predicts per‑Gaussian residuals for identity‑preserving, fine‑grained expression control, and a dual‑discriminator training scheme that enforces expression fidelity. The system achieves real‑time inference at 250 FPS on a single GPU and, for the first time, CPU‑only animatable 3DGS avatar synthesis at ~9 FPS.
By Ramazan Fazylov, Sergey Zagoruyko, Aleksandr Parkin, Stamatis Lefkimmiatis, Ivan Laptev
arXiv:2610.02207v1 Announce Type: cross
Abstract: 3D Gaussian avatars support fast rendering, however, their real-time animation is often challenged by the costly neural inference. We address this bo...
By Ramazan Fazylov, Stamatis Lefkimmiatis, Ivan Laptev
arXiv:2608. 19900v1 Announce Type: new Abstract: For full-body avatars, modeling surface dynamics is crucial for overcoming the uncanny valley and achieving perceptual realism.
By Guoxing Sun, Heming Zhu, Linjie Lyu, Pascal Fua, Christian Theobalt, Marc Habermann
arXiv:2609.18034v1 Announce Type: new
Abstract: Novel view synthesis from unposed multi-view images remains challenging, as the model must jointly learn scene representations and camera parameters wi...
By Wenyu Li, Sidun Liu, Peng Qiao, Yong Dou, Tongrui Hu
For full-body avatars, modeling surface dynamics is crucial for overcoming the uncanny valley and achieving perceptual realism. Person-agnostic methods recover static 3D avatars from monocular images, videos, or text prompts, but their skeleton-driven animations lack realistic surface dynamics such as clothing wrinkles.
PHOSA introduces MVSign, the first multi‑view Chinese sign language dataset co‑designed with Deaf experts, featuring diverse gestures and rich annotations. The authors develop a hybrid fitting pipeline for accurate SMPL‑X annotation and propose a decoupled sign avatar representation that isolates body, head, and hand components, coupled with a motion‑aware sampling strategy to handle motion blur and balance gesture diversity. Experiments show high‑fidelity visual results on MVSign, especially in detailed hand and facial regions, and good generalization to in‑the‑wild monocular sign language videos.
By Haodong Wang, Hezhen Hu, Wengang Zhou, Houqiang Li
AESplat is a new pose‑free feed‑forward 3D Gaussian Splatting framework that improves rendering quality by decoupling view‑independent and view‑dependent appearance modeling. It directly extracts the base view‑independent appearance from input images and predicts higher‑order spherical harmonic coefficients with a shallow MLP that incorporates 3D‑aware inductive biases. Experiments on several datasets show AESplat outperforms state‑of‑the‑art methods, achieving up to 0.8 dB higher PSNR than NAS3R and 1.1 dB over DepthSplat on RealEstate10K.
By Shiwei Ren, Zhiang Liu, Yongchun Fang, Hongwei Chen
arXiv:2609.38343v1 Announce Type: new
Abstract: We present SInGA, a novel method for learning Semantic Inpainting for animatable Gaussian head Avatars from a single image. Existing avatar approaches...
By Pilseo Park, Fizza Rubab, Yiying Tong
arXiv:2608.23410v1 Announce Type: new
Abstract: Photorealistic novel view synthesis of people remains challenging at high spatial resolutions and across multiple target cameras, where preserving iden...
By Federico Stella, Fei Jiang, Zhongshi Jiang, Zohar Barzelay, Emanuel Garbin, Amin Jourabloo, Liuhao Ge