paper-with-me

Papers

Talking-head Generation with Rhythmic Head Motion

2020-07-16 · Lele Chen, Guofeng Cui, Celong Liu, Zhong Li, Ziyi Kou, Yi Xu, Chenliang Xu

When people deliver a speech, they naturally move heads, and this rhythmic head motion conveys prosodic information. However, generating a lip-synced video while moving head naturally is challenging. While remarkably successful, existing works either generate still talkingface videos or rely on landmark/video frames as sparse/dense mapping guidance to generate head movements, which leads to unrealistic or uncontrollable video synthesis. To overcome the limitations, we propose a 3D-aware generative network along with a hybrid embedding module and a non-linear composition module. Through modeling the head motion and facial expressions1 explicitly, manipulating 3D animation carefully, and embedding reference images dynamically, our approach achieves controllable, photo-realistic, and temporally coherent talking-head videos with natural head movements. Thoughtful experiments on several standard benchmarks demonstrate that our method achieves significantly better results than the state-of-the-art methods in both quantitative and qualitative comparisons. The code is available on https://github.com/ lelechen63/Talking-head-Generation-with-Rhythmic-Head-Motion.

📄 PDF Abstract BibTeX arXiv:2007.08547

Code (1)

lelechen63/Talking-head-Generation-with-Rhythmic-Head-Motion 공식 구현 pytorch

Tasks

Talking Head Generation

Similar Papers 제목 키워드 기반

Write-a-speaker: Text-based Emotional and Rhythmic Talking-head Generation

2021-04-16 · Lincheng Li, Suzhen Wang, Zhimeng Zhang, Yu Ding 외

In this paper, we propose a novel text-based talking-head video generation framework that synthesizes high-fidelity facial expressions and head motions in accordance with contextual sentiments as well as speech rhythm an…

Face ModelRhythmTalking Head GenerationVideo Generation

OSM-Net: One-to-Many One-shot Talking Head Generation with Spontaneous Head Motions

2023-09-28 · Jin Liu, Xi Wang, Xiaomeng Fu, Yesheng Chai 외

One-shot talking head generation has no explicit head movement reference, thus it is difficult to generate talking heads with head motions. Some existing works only edit the mouth area and generate still talking heads, l…

Talking Head GenerationVideo Generation

Audio2Head: Audio-driven One-shot Talking-head Generation with Natural Head Motion

2021-07-20 · Suzhen Wang, Lincheng Li, Yu Ding, Changjie Fan 외

We propose an audio-driven talking-head method to generate photo-realistic talking-head videos from a single reference image. In this work, we tackle two key challenges: (i) producing natural head motions that match spee…

Image GenerationTalking Head Generation

A Keypoint Based Enhancement Method for Audio Driven Free View Talking Head Synthesis

2022-10-07 · Yichen Han, Ya Li, Yingming Gao, Jinlong Xue 외

Audio driven talking head synthesis is a challenging task that attracts increasing attention in recent years. Although existing methods based on 2D landmarks or 3D face models can synthesize accurate lip synchronization …

SPEAK: Speech-Driven Pose and Emotion-Adjustable Talking Head Generation

2024-05-12 · Changpeng Cai, Guinan Guo, Jiao Li, Junhao Su 외

Most earlier researches on talking face generation have focused on the synchronization of lip motion and speech content. However, head pose and facial emotions are equally important characteristics of natural faces. Whil…

DisentanglementFace GenerationTalking Face GenerationTalking Head Generation