paper-with-me

Papers

Controllable Talking Face Generation by Implicit Facial Keypoints Editing

2024-06-05 · Dong Zhao, Jiaying Shi, Wenjun Li, Shudong Wang, Shenghui Xu, Zhaoming Pan

Audio-driven talking face generation has garnered significant interest within the domain of digital human research. Existing methods are encumbered by intricate model architectures that are intricately dependent on each other, complicating the process of re-editing image or video inputs. In this work, we present ControlTalk, a talking face generation method to control face expression deformation based on driven audio, which can construct the head pose and facial expression including lip motion for both single image or sequential video inputs in a unified manner. By utilizing a pre-trained video synthesis renderer and proposing the lightweight adaptation, ControlTalk achieves precise and naturalistic lip synchronization while enabling quantitative control over mouth opening shape. Our experiments show that our method is superior to state-of-the-art performance on widely used benchmarks, including HDTF and MEAD. The parameterized adaptation demonstrates remarkable generalization capabilities, effectively handling expression deformation across same-ID and cross-ID scenarios, and extending its utility to out-of-domain portraits, regardless of languages. Code is available at https://github.com/NetEase-Media/ControlTalk.

📄 PDF Abstract BibTeX arXiv:2406.02880

Code (1)

NetEase-Media/ControlTalk 공식 구현 pytorch

Tasks

Face GenerationTalking Face Generation

Similar Papers 제목 키워드 기반

That's What I Said: Fully-Controllable Talking Face Generation

2023-04-06 · Youngjoon Jang, Kyeongha Rho, Jong-Bin Woo, Hyeongkeun Lee 외

The goal of this paper is to synthesise talking faces with controllable facial motions. To achieve this goal, we propose two key ideas. The first is to establish a canonical space where every face has the same motion pat…

Face GenerationNavigateTalking Face Generation

Dynamic Neural Textures: Generating Talking-Face Videos with Continuously Controllable Expressions

2022-04-13 · Zipeng Ye, Zhiyao Sun, Yu-Hui Wen, Yanan sun 외

Recently, talking-face video generation has received considerable attention. So far most methods generate results with neutral expressions or expressions that are implicitly determined by neural networks in an uncontroll…

Video Generation

Geometry-guided Emotion Modulation for Controllable and Photorealistic Emotional Talking Face Generation

2026-08-01 · Chenggong Hu, Shaoyin Ma, Yi Wang, Li Sun 외 arxiv

Audio-driven emotional talking face generation aims to synthesize realistic videos with expressive facial dynamics. However, existing methods struggle to balance controllability and visual fidelity. Although implicit rep…

Talking Face GenerationContinuous Control

OPT: One-shot Pose-Controllable Talking Head Generation

2023-02-16 · Jin Liu, Xi Wang, Xiaomeng Fu, Yesheng Chai 외

One-shot talking head generation produces lip-sync talking heads based on arbitrary audio and one source face. To guarantee the naturalness and realness, recent methods propose to achieve free pose control instead of sim…

DisentanglementTalking Head Generation

Playmate: Flexible Control of Portrait Animation via 3D-Implicit Space Guided Diffusion

2025-02-11 · Xingpei Ma, Jiaran Cai, Yuansheng Guan, Shenneng Huang 외

Recent diffusion-based talking face generation models have demonstrated impressive potential in synthesizing videos that accurately match a speech audio clip with a given reference identity. However, existing approaches …

AttributeDisentanglementFace GenerationPortrait Animation+1