paper-with-me

Papers

MotionMaster: Training-free Camera Motion Transfer For Video Generation

2024-04-24 · Teng Hu, Jiangning Zhang, Ran Yi, Yating Wang, Hongrui Huang, Jieyu Weng, Yabiao Wang, Lizhuang Ma

The emergence of diffusion models has greatly propelled the progress in image and video generation. Recently, some efforts have been made in controllable video generation, including text-to-video generation and video motion control, among which camera motion control is an important topic. However, existing camera motion control methods rely on training a temporal camera module, and necessitate substantial computation resources due to the large amount of parameters in video generation models. Moreover, existing methods pre-define camera motion types during training, which limits their flexibility in camera control. Therefore, to reduce training costs and achieve flexible camera control, we propose COMD, a novel training-free video motion transfer model, which disentangles camera motions and object motions in source videos and transfers the extracted camera motions to new videos. We first propose a one-shot camera motion disentanglement method to extract camera motion from a single source video, which separates the moving objects from the background and estimates the camera motion in the moving objects region based on the motion in the background by solving a Poisson equation. Furthermore, we propose a few-shot camera motion disentanglement method to extract the common camera motion from multiple videos with similar camera motions, which employs a window-based clustering technique to extract the common features in temporal attention maps of multiple videos. Finally, we propose a motion combination method to combine different types of camera motions together, enabling our model a more controllable and flexible camera control. Extensive experiments demonstrate that our training-free approach can effectively decouple camera-object motion and apply the decoupled camera motion to a wide range of controllable video generation tasks, achieving flexible and diverse camera motion control.

📄 PDF Abstract BibTeX arXiv:2404.15789

Code (0)

등록된 구현이 없습니다.

Tasks

DisentanglementMotion DisentanglementText-to-Video GenerationVideo Generation

Methods 이 논문이 사용한 방법론

Call To Westjet Airlines 설명 없음
Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Deep Homography Estimation in Dynamic Surgical Scenes for Laparoscopic Camera Motion Extraction

2021-09-30 · Martin Huber, Sébastien Ourselin, Christos Bergeles, Tom Vercauteren

Current laparoscopic camera motion automation relies on rule-based approaches or only focuses on surgical tools. Imitation Learning (IL) methods could alleviate these shortcomings, but have so far been applied to oversim…

CPUHomography EstimationImitation LearningMotion Estimation

MoRight: Motion Control Done Right

2026-04-08 · Shaowei Liu, Xuanchi Ren, Tianchang Shen, Huan Ling 외 arxiv

Generating motion-controlled videos--where user-specified actions drive physically plausible scene dynamics under freely chosen viewpoints--demands two capabilities: (1) disentangled motion control, allowing users to sep…

DreamCinema: Cinematic Transfer with Free Camera and 3D Character

2024-08-22 · Weiliang Chen, Fangfu Liu, Diankun Wu, Haowen Sun 외

We are living in a flourishing era of digital media, where everyone has the potential to become a personal filmmaker. Current research on cinematic transfer empowers filmmakers to reproduce and manipulate the visual elem…

MotionClone: Training-Free Motion Cloning for Controllable Video Generation

2024-06-08 · Pengyang Ling, Jiazi Bu, Pan Zhang, Xiaoyi Dong 외

Motion-based controllable video generation offers the potential for creating captivating visual content. Existing methods typically necessitate model training to encode particular motion cues or incorporate fine-tuning t…

DenoisingMotion GenerationMotion SynthesisText-to-Video Generation+1

MotionShop: Zero-Shot Motion Transfer in Video Diffusion Models with Mixture of Score Guidance

2024-12-06 · Hidir Yesiltepe, Tuna Han Salih Meral, Connor Dunlop, Pinar Yanardag

In this work, we propose the first motion transfer approach in diffusion transformer through Mixture of Score Guidance (MSG), a theoretically-grounded framework for motion transfer in diffusion models. Our key theoretica…

Object