paper-with-me

Papers

FLNet: Landmark Driven Fetching and Learning Network for Faithful Talking Facial Animation Synthesis

2019-11-21 · Kuangxiao Gu, Yuqian Zhou, Thomas Huang

Talking face synthesis has been widely studied in either appearance-based or warping-based methods. Previous works mostly utilize single face image as a source, and generate novel facial animations by merging other person's facial features. However, some facial regions like eyes or teeth, which may be hidden in the source image, can not be synthesized faithfully and stably. In this paper, We present a landmark driven two-stream network to generate faithful talking facial animation, in which more facial details are created, preserved and transferred from multiple source images instead of a single one. Specifically, we propose a network consisting of a learning and fetching stream. The fetching sub-net directly learns to attentively warp and merge facial regions from five source images of distinctive landmarks, while the learning pipeline renders facial organs from the training face space to compensate. Compared to baseline algorithms, extensive experiments demonstrate that the proposed method achieves a higher performance both quantitatively and qualitatively. Codes are at https://github.com/kgu3/FLNet_AAAI2020.

📄 PDF Abstract BibTeX arXiv:1911.09224

Code (0)

등록된 구현이 없습니다.

Tasks

Face Generation

Similar Papers 제목 키워드 기반

UniFLG: Unified Facial Landmark Generator from Text or Speech

2023-02-28 · Kentaro Mitsui, Yukiya Hono, Kei Sawada

Talking face generation has been extensively investigated owing to its wide applicability. The two primary frameworks used for talking face generation comprise a text-driven framework, which generates synchronized speech…

DecoderFace GenerationSpeech SynthesisTalking Face Generation+2

DreamHead: Learning Spatial-Temporal Correspondence via Hierarchical Diffusion for Audio-driven Talking Head Synthesis

2024-09-16 · Fa-Ting Hong, Yunfei Liu, Yu Li, Changyin Zhou 외

Audio-driven talking head synthesis strives to generate lifelike video portraits from provided audio. The diffusion model, recognized for its superior quality and robust generalization, has been explored for this task. H…

Talking Head Generation

EmoGene: Audio-Driven Emotional 3D Talking-Head Generation

2024-10-07 · Wenqing Wang, Yun Fu

Audio-driven talking-head generation is a crucial and useful technology for virtual human interaction and film-making. While recent advances have focused on improving image fidelity and lip synchronization, generating ac…

NeRFTalking Head Generation

Audio-driven High-resolution Seamless Talking Head Video Editing via StyleGAN

2024-07-08 · Jiacheng Su, KunHong Liu, Liyan Chen, Junfeng Yao 외

The existing methods for audio-driven talking head video editing have the limitations of poor visual effects. This paper tries to tackle this problem through editing talking face images seamless with different emotions b…

DisentanglementVideo Editing

KAN-Based Fusion of Dual-Domain for Audio-Driven Facial Landmarks Generation

2024-09-09 · Hoang-Son Vo-Thanh, Quang-Vinh Nguyen, Soo-Hyung Kim

Audio-driven talking face generation is a widely researched topic due to its high applicability. Reconstructing a talking face using audio significantly contributes to fields such as education, healthcare, online convers…

Face GenerationSpeech to Facial LandmarkTalking Face Generation