paper-with-me

Papers

Splat-Portrait: Generalizing Talking Heads with Gaussian Splatting

2026-01-26 · Tong Shi, Melonie de Almeida, Daniela Ivanova, Nicolas Pugeault, Paul Henderson arxiv

Talking Head Generation aims at synthesizing natural-looking talking videos from speech and a single portrait image. Previous 3D talking head generation methods have relied on domain-specific heuristics such as warping-based facial motion representation priors to animate talking motions, yet still produce inaccurate 3D avatar reconstructions, thus undermining the realism of generated animations. We introduce Splat-Portrait, a Gaussian-splatting-based method that addresses the challenges of 3D head reconstruction and lip motion synthesis. Our approach automatically learns to disentangle a single portrait image into a static 3D reconstruction represented as static Gaussian Splatting, and a predicted whole-image 2D background. It then generates natural lip motion conditioned on input audio, without any motion driven priors. Training is driven purely by 2D reconstruction and score-distillation losses, without 3D supervision nor landmarks. Experimental results demonstrate that Splat-Portrait exhibits superior performance on talking head generation and novel view synthesis, achieving better visual quality compared to previous works. Our project code and supplementary documents are public available at https://github.com/stonewalking/Splat-portrait.

📄 PDF Abstract BibTeX arXiv:2601.18633

Code (0)

등록된 구현이 없습니다.

Tasks

Talking Head GenerationNovel View Synthesis3D ReconstructionMotion Synthesis

Similar Papers 제목 키워드 기반

SyncTalk++: High-Fidelity and Efficient Synchronized Talking Heads Synthesis Using Gaussian Splatting

2025-06-17 · Ziqiao Peng, Wentao Hu, Junyuan Ma, Xiangyu Zhu 외

Achieving high synchronization in the synthesis of realistic, speech-driven talking head videos presents a significant challenge. A lifelike talking head requires synchronized coordination of subject identity, lip moveme…

GaussianHeadTalk: Wobble-Free 3D Talking Heads with Audio Driven Gaussian Splatting

2025-12-11 · Madhav Agarwal, Mingtian Zhang, Laura Sevilla-Lara, Steven McDonagh arxiv

Speech-driven talking heads have recently emerged and enable interactive avatars. However, real-world applications are limited, as current methods achieve high visual fidelity but slow or fast yet temporally unstable. Di…

Image Generation

DEGSTalk: Decomposed Per-Embedding Gaussian Fields for Hair-Preserving Talking Face Synthesis

2024-12-28 · Kaijun Deng, Dezhi Zheng, Jindong Xie, Jinbao Wang 외

Accurately synthesizing talking face videos and capturing fine facial features for individuals with long hair presents a significant challenge. To tackle these challenges in existing methods, we propose a decomposed per-…

3DGSFace Generation

TalkingGaussian: Structure-Persistent 3D Talking Head Synthesis via Gaussian Splatting

2024-04-23 · Jiahe Li, Jiawei Zhang, Xiao Bai, Jin Zheng 외

Radiance fields have demonstrated impressive performance in synthesizing lifelike 3D talking heads. However, due to the difficulty in fitting steep appearance changes, the prevailing paradigm that presents facial motions…

EmoTalkingGaussian: Continuous Emotion-conditioned Talking Head Synthesis

2025-02-02 · Junuk Cha, Seongro Yoon, Valeriya Strizhkova, Francois Bremond 외

3D Gaussian splatting-based talking head synthesis has recently gained attention for its ability to render high-fidelity images with real-time inference speed. However, since it is typically trained on only a short video…

Self-Supervised LearningSSIMtext-to-speechText to Speech