paper-with-me

Papers

X-Dyna: Expressive Dynamic Human Image Animation

2025-01-17 · CVPR 2025 1 · Di Chang, Hongyi Xu, You Xie, Yipeng Gao, Zhengfei Kuang, Shengqu Cai, Chenxu Zhang, Guoxian Song, Chao Wang, Yichun Shi, Zeyuan Chen, Shijie Zhou, Linjie Luo, Gordon Wetzstein, Mohammad Soleymani

We introduce X-Dyna, a novel zero-shot, diffusion-based pipeline for animating a single human image using facial expressions and body movements derived from a driving video, that generates realistic, context-aware dynamics for both the subject and the surrounding environment. Building on prior approaches centered on human pose control, X-Dyna addresses key shortcomings causing the loss of dynamic details, enhancing the lifelike qualities of human video animations. At the core of our approach is the Dynamics-Adapter, a lightweight module that effectively integrates reference appearance context into the spatial attentions of the diffusion backbone while preserving the capacity of motion modules in synthesizing fluid and intricate dynamic details. Beyond body pose control, we connect a local control module with our model to capture identity-disentangled facial expressions, facilitating accurate expression transfer for enhanced realism in animated scenes. Together, these components form a unified framework capable of learning physical human motion and natural scene dynamics from a diverse blend of human and scene videos. Comprehensive qualitative and quantitative evaluations demonstrate that X-Dyna outperforms state-of-the-art methods, creating highly lifelike and expressive animations. The code is available at https://github.com/bytedance/X-Dyna.

📄 PDF Abstract BibTeX arXiv:2501.10021

Code (1)

bytedance/x-dyna 공식 구현 pytorch

Tasks

Image Animation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Physically Plausible Animation of Human Upper Body from a Single Image

2022-12-09 · Ziyuan Huang, Zhengping Zhou, Yung-Yu Chuang, Jiajun Wu 외

We present a new method for generating controllable, dynamically responsive, and photorealistic human animations. Given an image of a person, our system allows the user to generate Physically plausible Upper Body Animati…

One-Shot Pose-Driving Face Animation Platform

2024-07-12 · He Feng, Donglin Di, Yongjia Ma, Wei Chen 외

The objective of face animation is to generate dynamic and expressive talking head videos from a single reference face, utilizing driving conditions derived from either video or audio inputs. Current approaches often req…

HairWeaver: Few-Shot Photorealistic Hair Motion Synthesis with Sim-to-Real Guided Video Diffusion

2026-02-11 · Di Chang, Ji Hou, Aljaz Bozic, Assaf Neuberger 외 arxiv

We present HairWeaver, a diffusion-based pipeline that animates a single human image with realistic and expressive hair dynamics. While existing methods successfully control body pose, they lack specific control over hai…

Motion Synthesis

ReFree: Towards Realistic Co-Speech Video Generation via Reward-Free RL and Multilevel Speech Guidance

2026-06-11 · Salaheldin Mohamed, M. Hamza Mughal, Rishabh Dabral, Christian Theobalt arxiv

Speech-driven talking character animation seeks to generate life-like portrait videos that convey natural conversation behavior, aligning facial motion with spoken audio. Although recent advances in video generation have…

Reinforcement LearningVideo Generation

EchoMimicV2: Towards Striking, Simplified, and Semi-Body Human Animation

2024-11-15 · CVPR 2025 1 · Rang Meng, Xingyu Zhang, Yuming Li, Chenguang Ma

Recent work on human animation usually involves audio, pose, or movement maps conditions, thereby achieves vivid animation quality. However, these methods often face practical challenges due to extra control conditions, …

Audio-Driven Body AnimationHuman AnimationImage to Video GenerationSubject-driven Video Generation+1