paper-with-me

Papers

EverAnimate: Minute-Scale Human Animation via Latent Flow Restoration

2026-05-14 · Wuyang Li, Yang Gao, Mariam Hassan, Lan Feng, Wentao Pan, Po-Chien Luan, Alexandre Alahi arxiv

We propose EverAnimate, an efficient post-training method for long-horizon animated video generation that preserves visual quality and character identity. Long-form animation remains challenging because highly dynamic human motion must be synthesized against relatively static environments, making chunk-based generation prone to accumulated drift: (i) low-level quality drift, such as progressive degradation of static backgrounds, and (ii) high-level semantic drift, such as inconsistent character identity and view-dependent attributes. To address this issue, EverAnimate restores drifted flow trajectories by anchoring generation to a persistent latent context memory, consisting of two complementary mechanisms. (i) Persistent Latent Propagation maintains a context memory across chunks to propagate identity and motion in latent space while mitigating temporal forgetting. (ii) Restorative Flow Matching introduces an implicit restoration objective during sampling through velocity adjustment, improving within-chunk fidelity. With only lightweight LoRA tuning, EverAnimate outperforms state-of-the-art long-animation methods in both short- and long-horizon settings: at 10 seconds, it improves PSNR/SSIM by 8%/7% and reduces LPIPS/FID by 22%/11%; at 90 seconds, the gains increase to 15%/15% and 32%/27%, respectively.

📄 PDF Abstract BibTeX arXiv:2605.15042

Code (0)

등록된 구현이 없습니다.

Tasks

Video Generation

Similar Papers 제목 키워드 기반

LiveAnimate: Stable Long-Form Streaming Human Animation in Real-Time

2026-08-12 · Yuxuan Zhang, Haozhong Xiong, Yubo Huang, Jiayi Song 외 arxiv

Pose-driven human animation synthesizes a video of a target person from a single reference image and a driving pose stream. Real-time generation is essential for interactive applications such as live streaming, teleprese…

Animating Petascale Time-varying Data on Commodity Hardware with LLM-assisted Scripting

2026-03-07 · Ishrat Jahan Eliza, Xuan Huang, Aashish Panta, Alper Sahistan 외 arxiv

Scientists face significant visualization challenges as time-varying datasets grow in speed and volume, often requiring specialized infrastructure and expertise to handle massive datasets. Petascale climate models genera…

Audio-Driven Facial Animation by Joint End-to-End Learning of Pose and Emotion

2017-07-30 · SIGGRAPH 2017 7 · Tero Karras, Timo Aila, Samuli Laine, Antti Herva 외

We present a machine learning technique for driving 3D facial animation by audio input in real time and with low latency. Our deep neural network learns a mapping from input waveforms to the 3D vertex coordinates of a fa…

Face Model

Human Geometry Distribution for 3D Animation Generation

2025-12-08 · Xiangjun Tang, Biao Zhang, Peter Wonka arxiv

Generating realistic human geometry animations remains a challenging task, as it requires modeling natural clothing dynamics with fine-grained geometric details under limited data. To address these challenges, we propose…

SoulX-LiveAct: Towards Hour-Scale Real-Time Human Animation with Neighbor Forcing and ConvKV Memory

2026-03-12 · Dingcheng Zhen, Xu Zheng, Ruixin Zhang, Zhiqi Jiang 외 arxiv

Autoregressive (AR) diffusion models offer a promising framework for sequential generation tasks such as video synthesis by combining diffusion modeling with causal inference. Although they support streaming generation, …

Causal InferenceVideo Generation