paper-with-me

홈 › Papers

MIMAFace: Face Animation via Motion-Identity Modulated Appearance Feature Learning

2024-09-23 · Yue Han, Junwei Zhu, Yuxiang Feng, Xiaozhong Ji, Keke He, Xiangtai Li, Zhucun Xue, Yong liu

Current diffusion-based face animation methods generally adopt a ReferenceNet (a copy of U-Net) and a large amount of curated self-acquired data to learn appearance features, as robust appearance features are vital for ensuring temporal stability. However, when trained on public datasets, the results often exhibit a noticeable performance gap in image quality and temporal consistency. To address this issue, we meticulously examine the essential appearance features in the facial animation tasks, which include motion-agnostic (e.g., clothing, background) and motion-related (e.g., facial details) texture components, along with high-level discriminative identity features. Drawing from this analysis, we introduce a Motion-Identity Modulated Appearance Learning Module (MIA) that modulates CLIP features at both motion and identity levels. Additionally, to tackle the semantic/ color discontinuities between clips, we design an Inter-clip Affinity Learning Module (ICA) to model temporal relationships across clips. Our method achieves precise facial motion control (i.e., expressions and gaze), faithful identity preservation, and generates animation videos that maintain both intra/inter-clip temporal consistency. Moreover, it easily adapts to various modalities of driving sources. Extensive experiments demonstrate the superiority of our method.

📄 PDF Abstract BibTeX arXiv:2409.15179

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

CLIP Contrastive Language-Image Pre-training (CLIP), consisting of a simplified version of ConVIRT trained from scratch, is an efficient method of image representation learning…

Similar Papers 제목 키워드 기반

Synthesis for the Kinematic Control of Identity in Sign Language

2022-06-01 · SLTAT (LREC) 2022 6 · Félix Bigand, Elise Prigent, Annelies Braffort

Sign Language (SL) animations generated from motion capture (mocap) of real signers convey critical information about their identity. It has been suggested that this information is mostly carried by statistics of the mov…

DF-3DFace: One-to-Many Speech Synchronized 3D Face Animation with Diffusion

2023-08-23 · Se Jin Park, Joanna Hong, Minsu Kim, Yong Man Ro

Speech-driven 3D facial animation has gained significant attention for its ability to create realistic and expressive facial animations in 3D space based on speech. Learning-based methods have shown promising progress in…

3D Face Animation

Learning Semantic Facial Descriptors for Accurate Face Animation

2025-01-29 · Lei Zhu, Yuanqi Chen, Xiaohang Liu, Thomas H. Li 외

Face animation is a challenging task. Existing model-based methods (utilizing 3DMMs or landmarks) often result in a model-like reconstruction effect, which doesn't effectively preserve identity. Conversely, model-free ap…

LSF-Animation: Label-Free Speech-Driven Facial Animation via Implicit Feature Representation

2025-10-23 · Xin Lu, Chuanqing Zhuang, Chenxi Jin, Zhengda Lu 외 arxiv

Speech-driven 3D facial animation has attracted increasing interest since its potential to generate expressive and temporally synchronized digital humans. While recent works have begun to explore emotion-aware animation,…

Single Source One Shot Reenactment using Weighted motion From Paired Feature Points

2021-04-07 · Soumya Tripathy, Juho Kannala, Esa Rahtu

Image reenactment is a task where the target object in the source image imitates the motion represented in the driving image. One of the most common reenactment tasks is face image animation. The major challenge in the c…

Face ReenactmentImage Animation