paper-with-me

홈 › Papers

Generative Adversarial Talking Head: Bringing Portraits to Life with a Weakly Supervised Neural Network

2018-03-21 · Hai X. Pham, Yuting Wang, Vladimir Pavlovic

This paper presents Generative Adversarial Talking Head (GATH), a novel deep generative neural network that enables fully automatic facial expression synthesis of an arbitrary portrait with continuous action unit (AU) coefficients. Specifically, our model directly manipulates image pixels to make the unseen subject in the still photo express various emotions controlled by values of facial AU coefficients, while maintaining her personal characteristics, such as facial geometry, skin color and hair style, as well as the original surrounding background. In contrast to prior work, GATH is purely data-driven and it requires neither a statistical face model nor image processing tricks to enact facial deformations. Additionally, our model is trained from unpaired data, where the input image, with its auxiliary identity label taken from abundance of still photos in the wild, and the target frame are from different persons. In order to effectively learn such model, we propose a novel weakly supervised adversarial learning framework that consists of a generator, a discriminator, a classifier and an action unit estimator. Our work gives rise to template-and-target-free expression editing, where still faces can be effortlessly animated with arbitrary AU coefficients provided by the user.

📄 PDF Abstract BibTeX arXiv:1803.07716

Code (0)

등록된 구현이 없습니다.

Tasks

Face Model

Similar Papers 제목 키워드 기반

Silence is Golden: Leveraging Adversarial Examples to Nullify Audio Control in LDM-based Talking-Head Generation

2025-06-02 · CVPR 2025 1 · Yuan Gan, Jiaxu Miao, Yunze Wang, Yi Yang

Advances in talking-head animation based on Latent Diffusion Models (LDM) enable the creation of highly realistic, synchronized videos. These fabricated videos are indistinguishable from real ones, increasing the risk of…

MisinformationTalking Head Generation

GMTalker: Gaussian Mixture-based Audio-Driven Emotional Talking Video Portraits

2023-12-12 · Yibo Xia, Lizhen Wang, Xiang Deng, Xiaoyan Luo 외

Synthesizing high-fidelity and emotion-controllable talking video portraits, with audio-lip sync, vivid expressions, realistic head poses, and eye blinks, has been an important and challenging task in recent years. Most …

Diversity

Real-time Neural Radiance Talking Portrait Synthesis via Audio-spatial Decomposition

2022-11-22 · Jiaxiang Tang, Kaisiyuan Wang, Hang Zhou, Xiaokang Chen 외

While dynamic Neural Radiance Fields (NeRF) have shown success in high-fidelity 3D modeling of talking portraits, the slow training and inference speed severely obstruct their potential usage. In this paper, we propose a…

NeRFTalking Face Generation

EmoGene: Audio-Driven Emotional 3D Talking-Head Generation

2024-10-07 · Wenqing Wang, Yun Fu

Audio-driven talking-head generation is a crucial and useful technology for virtual human interaction and film-making. While recent advances have focused on improving image fidelity and lip synchronization, generating ac…

NeRFTalking Head Generation

From Pixels to Portraits: A Comprehensive Survey of Talking Head Generation Techniques and Applications

2023-08-30 · Shreyank N Gowda, Dheeraj Pandey, Shashank Narayana Gowda

Recent advancements in deep learning and computer vision have led to a surge of interest in generating realistic talking heads. This paper presents a comprehensive survey of state-of-the-art methods for talking head gene…

NeRFSurveyTalking Head Generation