paper-with-me

Papers

3D-Aware Talking-Head Video Motion Transfer

2023-11-05 · Haomiao Ni, Jiachen Liu, Yuan Xue, Sharon X. Huang

Motion transfer of talking-head videos involves generating a new video with the appearance of a subject video and the motion pattern of a driving video. Current methodologies primarily depend on a limited number of subject images and 2D representations, thereby neglecting to fully utilize the multi-view appearance features inherent in the subject video. In this paper, we propose a novel 3D-aware talking-head video motion transfer network, Head3D, which fully exploits the subject appearance information by generating a visually-interpretable 3D canonical head from the 2D subject frames with a recurrent network. A key component of our approach is a self-supervised 3D head geometry learning module, designed to predict head poses and depth maps from 2D subject video frames. This module facilitates the estimation of a 3D head in canonical space, which can then be transformed to align with driving video frames. Additionally, we employ an attention-based fusion network to combine the background and other details from subject frames with the 3D subject head to produce the synthetic target video. Our extensive experiments on two public talking-head video datasets demonstrate that Head3D outperforms both 2D and 3D prior arts in the practical cross-identity setting, with evidence showing it can be readily adapted to the pose-controllable novel view synthesis task.

📄 PDF Abstract BibTeX arXiv:2311.02549

Code (0)

등록된 구현이 없습니다.

Tasks

Novel View Synthesis

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

High-Fidelity and Freely Controllable Talking Head Video Generation

2023-04-20 · CVPR 2023 1 · Yue Gao, Yuan Zhou, Jinglu Wang, Xiao Li 외

Talking head generation is to generate video based on a given source identity and target motion. However, current methods face several challenges that limit the quality and controllability of the generated videos. First,…

Face ModelTalking Head GenerationVideo GenerationVocal Bursts Intensity Prediction

Audio2Head: Audio-driven One-shot Talking-head Generation with Natural Head Motion

2021-07-20 · Suzhen Wang, Lincheng Li, Yu Ding, Changjie Fan 외

We propose an audio-driven talking-head method to generate photo-realistic talking-head videos from a single reference image. In this work, we tackle two key challenges: (i) producing natural head motions that match spee…

Image GenerationTalking Head Generation

Efficient Emotional Adaptation for Audio-Driven Talking-Head Generation

2023-09-10 · ICCV 2023 1 · Yuan Gan, Zongxin Yang, Xihang Yue, Lingyun Sun 외

Audio-driven talking-head synthesis is a popular research topic for virtual human-related applications. However, the inflexibility and inefficiency of existing methods, which necessitate expensive end-to-end training to …

Talking Head Generation

Talking-head Generation with Rhythmic Head Motion

2020-07-16 · Lele Chen, Guofeng Cui, Celong Liu, Zhong Li 외

When people deliver a speech, they naturally move heads, and this rhythmic head motion conveys prosodic information. However, generating a lip-synced video while moving head naturally is challenging. While remarkably suc…

Talking Head Generation

EmoCAST: Emotional Talking Portrait via Emotive Text Description

2025-08-28 · Yiguo Jiang, Xiaodong Cun, Yong Zhang, Yudian Zheng 외 arxiv

Emotional talking head synthesis aims to generate talking portrait videos with vivid expressions. Existing methods still exhibit limitations in control flexibility, motion naturalness, and expression quality. Moreover, c…

Motion Synthesis