paper-with-me

Papers

CP-EB: Talking Face Generation with Controllable Pose and Eye Blinking Embedding

2023-11-15 · Jianzong Wang, Yimin Deng, ZiQi Liang, xulong Zhang, Ning Cheng, Jing Xiao

This paper proposes a talking face generation method named "CP-EB" that takes an audio signal as input and a person image as reference, to synthesize a photo-realistic people talking video with head poses controlled by a short video clip and proper eye blinking embedding. It's noted that not only the head pose but also eye blinking are both important aspects for deep fake detection. The implicit control of poses by video has already achieved by the state-of-art work. According to recent research, eye blinking has weak correlation with input audio which means eye blinks extraction from audio and generation are possible. Hence, we propose a GAN-based architecture to extract eye blink feature from input audio and reference video respectively and employ contrastive training between them, then embed it into the concatenated features of identity and poses to generate talking face images. Experimental results show that the proposed method can generate photo-realistic talking face with synchronous lips motions, natural head poses and blinking eyes.

📄 PDF Abstract BibTeX arXiv:2311.08673

Code (0)

등록된 구현이 없습니다.

Tasks

Face GenerationTalking Face Generation

Methods 이 논문이 사용한 방법론

CLIP Contrastive Language-Image Pre-training (CLIP), consisting of a simplified version of ConVIRT trained from scratch, is an efficient method of image representation learning…

Similar Papers 제목 키워드 기반

MoDiT: Learning Highly Consistent 3D Motion Coefficients with Diffusion Transformer for Talking Head Generation

2025-07-07 · Yucheng Wang, Dan Xu arxiv

Audio-driven talking head generation is critical for applications such as virtual assistants, video games, and films, where natural lip movements are essential. Despite progress in this field, challenges remain in produc…

Talking Head GenerationInformation Extraction

OPT: One-shot Pose-Controllable Talking Head Generation

2023-02-16 · Jin Liu, Xi Wang, Xiaomeng Fu, Yesheng Chai 외

One-shot talking head generation produces lip-sync talking heads based on arbitrary audio and one source face. To guarantee the naturalness and realness, recent methods propose to achieve free pose control instead of sim…

DisentanglementTalking Head Generation

That's What I Said: Fully-Controllable Talking Face Generation

2023-04-06 · Youngjoon Jang, Kyeongha Rho, Jong-Bin Woo, Hyeongkeun Lee 외

The goal of this paper is to synthesise talking faces with controllable facial motions. To achieve this goal, we propose two key ideas. The first is to establish a canonical space where every face has the same motion pat…

Face GenerationNavigateTalking Face Generation

Taming Transformer for Emotion-Controllable Talking Face Generation

2025-08-20 · Ziqi Zhang, Cheng Deng arxiv

Talking face generation is a novel and challenging generation task, aiming at synthesizing a vivid speaking-face video given a specific audio. To fulfill emotion-controllable talking face generation, current methods need…

Talking Face Generation

Towards Flexible, Natural, Efficient Interaction for Conversational Talking Face Generation

2026-06-30 · Baiqin Wang, Sen Chen, Jiankuo Zhao, Xiangyu Liu 외 arxiv

Conversational talking face generation has recently attracted increasing attention, aiming to synthesize interactive talking videos where characters speak, listen, and respond dynamically to each other. This task present…

Talking Face GenerationData Augmentation