arXiv AI By Ziheng Yang, Yinfeng Yu, Yongming Li

Talking Head Synthesis with Facial Landmark Guidance via 3D Gaussian Splatting

Read the original on arXiv AI →

The Flow has not summarised this story yet — read it at arXiv AI.

arXiv Computer Vision
Sep 4

EmbedTalk: Talking Head Synthesis using Gaussian Embeddings

EmbedTalk introduces per‑Gaussian embeddings to drive speech‑driven facial deformations in real‑time talking head synthesis, replacing traditional tri‑plane encodings. This approach improves rendering quality, lip synchronisation, and motion consistency compared to prior 3D Gaussian Splatting methods while producing more compact models that run at 60+ FPS on a laptop GPU. The technique demonstrates competitive performance against state‑of‑the‑art generative models.

By Arpita Saggar, Jonathan C. Darling, Duygu Sarikaya, David C. Hogg
Hugging Face Trending Papers
Jul 1

GaussianEmoTalker: Real-Time Emotional Talking Head Synthesis with Audio-Driven and Blendshape-Based 3D Gaussian Splatting

Audio-driven talking head synthesis has achieved impressive progress in lip synchronization and visual quality, yet generating expressive emotional avatars with controllable intensity remains challenging, especially under real-time constraints. In this paper, we present GaussianEmoTalker, an audio-driven framework for real-time emotional talking head synthesis based on 3D Gaussian Splatting.