paper-with-me

Papers

Embedded Representation Learning Network for Animating Styled Video Portrait

2024-04-29 · Tianyong Wang, Xiangyu Liang, Wangguandong Zheng, Dan Niu, Haifeng Xia, Siyu Xia

The talking head generation recently attracted considerable attention due to its widespread application prospects, especially for digital avatars and 3D animation design. Inspired by this practical demand, several works explored Neural Radiance Fields (NeRF) to synthesize the talking heads. However, these methods based on NeRF face two challenges: (1) Difficulty in generating style-controllable talking heads. (2) Displacement artifacts around the neck in rendered images. To overcome these two challenges, we propose a novel generative paradigm \textit{Embedded Representation Learning Network} (ERLNet) with two learning stages. First, the \textit{ audio-driven FLAME} (ADF) module is constructed to produce facial expression and head pose sequences synchronized with content audio and style video. Second, given the sequence deduced by the ADF, one novel \textit{dual-branch fusion NeRF} (DBF-NeRF) explores these contents to render the final images. Extensive empirical studies demonstrate that the collaboration of these two stages effectively facilitates our method to render a more realistic talking head than the existing algorithms.

📄 PDF Abstract BibTeX arXiv:2404.19038

Code (0)

등록된 구현이 없습니다.

Tasks

NeRFRepresentation LearningTalking Head Generation

Similar Papers 제목 키워드 기반

MODA: Mapping-Once Audio-driven Portrait Animation with Dual Attentions

2023-07-19 · ICCV 2023 1 · Yunfei Liu, Lijian Lin, Fei Yu, Changyin Zhou 외

Audio-driven portrait animation aims to synthesize portrait videos that are conditioned by given audio. Animating high-fidelity and multimodal video portraits has a variety of applications. Previous methods have attempte…

Portrait Animation

PV3D: A 3D Generative Model for Portrait Video Generation

2022-12-13 · Zhongcong Xu, Jianfeng Zhang, Jun Hao Liew, Wenqing Zhang 외

Recent advances in generative adversarial networks (GANs) have demonstrated the capabilities of generating stunning photo-realistic portrait images. While some prior works have applied such image GANs to unconditional 2D…

Video Generation

Semantic-Aware Implicit Neural Audio-Driven Video Portrait Generation

2022-01-19 · Xian Liu, Yinghao Xu, Qianyi Wu, Hang Zhou 외

Animating high-fidelity video portrait with speech audio is crucial for virtual reality and digital entertainment. While most previous studies rely on accurate explicit structural information, recent works explore the im…

NeRF

Hallo3: Highly Dynamic and Realistic Portrait Image Animation with Video Diffusion Transformer

2024-12-01 · CVPR 2025 1 · Jiahao Cui, Hui Li, Yun Zhan, Hanlin Shang 외

Existing methodologies for animating portrait images face significant challenges, particularly in handling non-frontal perspectives, rendering dynamic objects around the portrait, and generating immersive, realistic back…

Image AnimationPortrait Animation

SPACE: Speech-driven Portrait Animation with Controllable Expression

2022-11-17 · ICCV 2023 1 · Siddharth Gururani, Arun Mallya, Ting-Chun Wang, Rafael Valle 외

Animating portraits using speech has received growing attention in recent years, with various creative and practical use cases. An ideal generated video should have good lip sync with the audio, natural facial expression…

Portrait Animation