paper-with-me

Papers

Cinemo: Consistent and Controllable Image Animation with Motion Diffusion Models

2024-07-22 · Xin Ma, Yaohui Wang, Gengyun Jia, Xinyuan Chen, Yuan-Fang Li, Cunjian Chen, Yu Qiao

Diffusion models have achieved great progress in image animation due to powerful generative capabilities. However, maintaining spatio-temporal consistency with detailed information from the input static image over time (e.g., style, background, and object of the input static image) and ensuring smoothness in animated video narratives guided by textual prompts still remains challenging. In this paper, we introduce Cinemo, a novel image animation approach towards achieving better motion controllability, as well as stronger temporal consistency and smoothness. In general, we propose three effective strategies at the training and inference stages of Cinemo to accomplish our goal. At the training stage, Cinemo focuses on learning the distribution of motion residuals, rather than directly predicting subsequent via a motion diffusion model. Additionally, a structural similarity index-based strategy is proposed to enable Cinemo to have better controllability of motion intensity. At the inference stage, a noise refinement technique based on discrete cosine transformation is introduced to mitigate sudden motion changes. Such three strategies enable Cinemo to produce highly consistent, smooth, and motion-controllable results. Compared to previous methods, Cinemo offers simpler and more precise user controllability. Extensive experiments against several state-of-the-art methods, including both commercial tools and research approaches, across multiple metrics, demonstrate the effectiveness and superiority of our proposed approach.

📄 PDF Abstract BibTeX arXiv:2407.15642

Code (1)

maxin-cn/Cinemo 공식 구현 pytorch

Tasks

Image Animation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Consistent and Controllable Image Animation with Motion Diffusion Models

2025-01-01 · CVPR 2025 1 · Xin Ma, Yaohui Wang, Gengyun Jia, Xinyuan Chen 외

Diffusion models have achieved significant progress in the task of image animation due to their powerful generative capabilities. However, preserving appearance consistency to the static input image, and avoiding abr…

Image AnimationVideo Editing

CineMobile: On-Device Image-to-Video Diffusion for Cinematic Camera Motion Generation

2026-07-04 · Xuyao Huang, Zelai Deng, Xu Wang, Xizhong Xiao 외 hf

The growing demand for image-to-video creation on mobile devices has increasingly focused on cinematic motion effects like bullet time, dolly zoom, slow motion, etc. While Diffusion Transformers (DiTs) exhibit strong per…

Reinforcement LearningVideo Generation

Consistent and Controllable Image Animation with Motion Linear Diffusion Transformers

2025-08-10 · Xin Ma, Yaohui Wang, Genyun Jia, Xinyuan Chen 외 arxiv

Image animation has seen significant progress, driven by the powerful generative capabilities of diffusion models. However, maintaining appearance consistency with static input images and mitigating abrupt motion transit…

Perception-as-Control: Fine-grained Controllable Image Animation with 3D-aware Motion Representation

2025-01-09 · Yingjie Chen, Yifang Men, Yuan YAO, Miaomiao Cui 외

Motion-controllable image animation is a fundamental task with a wide range of potential applications. Recent works have made progress in controlling camera or object motion via various motion representations, while they…

Image AnimationObject

CharacterShot: Controllable and Consistent 4D Character Animation

2025-08-10 · Junyao Gao, Jiaxing Li, Wenran Liu, Yanhong Zeng 외 arxiv

In this paper, we propose \textbf{CharacterShot}, a controllable and consistent 4D character animation framework that enables any individual designer to create dynamic 3D characters (i.e., 4D character animation) from a …