paper-with-me

Papers

AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

2024-03-26 · Huawei Wei, Zejun Yang, Zhisheng Wang

In this study, we propose AniPortrait, a novel framework for generating high-quality animation driven by audio and a reference portrait image. Our methodology is divided into two stages. Initially, we extract 3D intermediate representations from audio and project them into a sequence of 2D facial landmarks. Subsequently, we employ a robust diffusion model, coupled with a motion module, to convert the landmark sequence into photorealistic and temporally consistent portrait animation. Experimental results demonstrate the superiority of AniPortrait in terms of facial naturalness, pose diversity, and visual quality, thereby offering an enhanced perceptual experience. Moreover, our methodology exhibits considerable potential in terms of flexibility and controllability, which can be effectively applied in areas such as facial motion editing or face reenactment. We release code and model weights at https://github.com/scutzzj/AniPortrait

📄 PDF Abstract BibTeX arXiv:2403.17694

Code (2)

scutzzj/aniportrait 공식 구현 pytorch
zejun-yang/aniportrait pytorch

Tasks

DiversityFace ReenactmentPortrait Animation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

LinguaLinker: Audio-Driven Portraits Animation with Implicit Facial Control Enhancement

2024-07-26 · Rui Zhang, Yixiao Fang, Zhengnan Lu, Pei Cheng 외

This study delves into the intricacies of synchronizing facial dynamics with multilingual audio inputs, focusing on the creation of visually compelling, time-synchronized animations through diffusion-based techniques. Di…

Hallo4: High-Fidelity Dynamic Portrait Animation via Direct Preference Optimization and Temporal Motion Modulation

2025-05-29 · Jiahao Cui, Yan Chen, Mingwang Xu, Hanlin Shang 외

Generating highly dynamic and photorealistic portrait animations driven by audio and skeletal motion remains challenging due to the need for precise lip synchronization, natural facial expressions, and high-fidelity body…

Portrait AnimationVideo Alignment

FREAK: Frequency-modulated High-fidelity and Real-time Audio-driven Talking Portrait Synthesis

2025-03-06 · Ziqi Ni, Ao Fu, Yi Zhou

Achieving high-fidelity lip-speech synchronization in audio-driven talking portrait synthesis remains challenging. While multi-stage pipelines or diffusion models yield high-quality results, they suffer from high computa…

Audio-Visual Synchronization

Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

2024-06-13 · Mingwang Xu, Hui Li, Qingkun Su, Hanlin Shang 외

The field of portrait image animation, driven by speech audio input, has experienced significant advancements in the generation of realistic and dynamic portraits. This research delves into the complexities of synchroniz…

DiversityImage Animation

Loopy: Taming Audio-Driven Portrait Avatar with Long-Term Motion Dependency

2024-09-04 · Jianwen Jiang, Chao Liang, Jiaqi Yang, Gaojie Lin 외

With the introduction of diffusion-based video generation techniques, audio-conditioned human video generation has recently achieved significant breakthroughs in both the naturalness of motion and the synthesis of portra…

Video Generation