paper-with-me

홈 › Papers

I2VControl: Disentangled and Unified Video Motion Synthesis Control

2024-11-26 · Wanquan Feng, Tianhao Qi, Jiawei Liu, Mingzhen Sun, Pengqi Tu, Tianxiang Ma, Fei Dai, Songtao Zhao, Siyu Zhou, Qian He

Video synthesis techniques are undergoing rapid progress, with controllability being a significant aspect of practical usability for end-users. Although text condition is an effective way to guide video synthesis, capturing the correct joint distribution between text descriptions and video motion remains a substantial challenge. In this paper, we present a disentangled and unified framework, namely I2VControl, that unifies multiple motion control tasks in image-to-video synthesis. Our approach partitions the video into individual motion units and represents each unit with disentangled control signals, which allows for various control types to be flexibly combined within our single system. Furthermore, our methodology seamlessly integrates as a plug-in for pre-trained models and remains agnostic to specific model architectures. We conduct extensive experiments, achieving excellent performance on various control tasks, and our method further facilitates user-driven creative combinations, enhancing innovation and creativity. The project page is: https://wanquanf.github.io/I2VControl .

📄 PDF Abstract BibTeX arXiv:2411.17765

Code (0)

등록된 구현이 없습니다.

Tasks

Motion Synthesis

Similar Papers 제목 키워드 기반

Progressive Disentangled Representation Learning for Fine-Grained Controllable Talking Head Synthesis

2022-11-26 · CVPR 2023 1 · Duomin Wang, Yu Deng, Zixin Yin, Heung-Yeung Shum 외

We present a novel one-shot talking head synthesis method that achieves disentangled and fine-grained control over lip motion, eye gaze&blink, head pose, and emotional expression. We represent different motions via disen…

Contrastive LearningDisentanglementRepresentation Learning

I2VControl-Camera: Precise Video Camera Control with Adjustable Motion Strength

2024-11-10 · Wanquan Feng, Jiawei Liu, Pengqi Tu, Tianhao Qi 외

Video generation technologies are developing rapidly and have broad potential applications. Among these technologies, camera control is crucial for generating professional-quality videos that accurately meet user expecta…

Video Generation

DivControl: Knowledge Diversion for Controllable Image Generation

2025-07-31 · Yucheng Xie, Fu Feng, Ruixiao Shi, Jing Wang 외 arxiv

Diffusion models have advanced from text-to-image (T2I) to image-to-image (I2I) generation by incorporating structured inputs such as depth maps, enabling fine-grained spatial control. However, existing methods either tr…

Zero-shot GeneralizationImage Generation

DEMO: Disentangled Motion Latent Flow Matching for Fine-Grained Controllable Talking Portrait Synthesis

2025-10-12 · Peiyin Chen, Zhuowei Yang, Hui Feng, Sheng Jiang 외 arxiv

Audio-driven talking-head generation has advanced rapidly with diffusion-based generative models, yet producing temporally coherent videos with fine-grained motion control remains challenging. We propose DEMO, a flow-mat…

Disentangled Motion Modeling for Video Frame Interpolation

2024-06-25 · Jaihyun Lew, Jooyoung Choi, Chaehun Shin, Dahuin Jung 외

Video Frame Interpolation (VFI) aims to synthesize intermediate frames between existing frames to enhance visual smoothness and quality. Beyond the conventional methods based on the reconstruction loss, recent works have…

Optical Flow EstimationVideo Frame Interpolation