paper-with-me

Papers

SG-I2V: Self-Guided Trajectory Control in Image-to-Video Generation

2024-11-07 · Koichi Namekata, Sherwin Bahmani, Ziyi Wu, Yash Kant, Igor Gilitschenski, David B. Lindell

Methods for image-to-video generation have achieved impressive, photo-realistic quality. However, adjusting specific elements in generated videos, such as object motion or camera movement, is often a tedious process of trial and error, e.g., involving re-generating videos with different random seeds. Recent techniques address this issue by fine-tuning a pre-trained model to follow conditioning signals, such as bounding boxes or point trajectories. Yet, this fine-tuning procedure can be computationally expensive, and it requires datasets with annotated object motion, which can be difficult to procure. In this work, we introduce SG-I2V, a framework for controllable image-to-video generation that is self-guided$\unicode{x2013}$offering zero-shot control by relying solely on the knowledge present in a pre-trained image-to-video diffusion model without the need for fine-tuning or external knowledge. Our zero-shot method outperforms unsupervised baselines while significantly narrowing down the performance gap with supervised models in terms of visual quality and motion fidelity.

📄 PDF Abstract BibTeX arXiv:2411.04989

Code (0)

등록된 구현이 없습니다.

Tasks

Image to Video GenerationVideo Generation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Trajectory Attention for Fine-grained Video Motion Control

2024-11-28 · Zeqi Xiao, Wenqi Ouyang, Yifan Zhou, Shuai Yang 외

Recent advancements in video generation have been greatly driven by video diffusion models, with camera motion control emerging as a crucial challenge in creating view-customized visual content. This paper introduces tra…

Inductive BiasVideo EditingVideo Generation

Zo3T: Zero-Shot 3D-Aware Trajectory-Guided Image-to-Video Generation via Test-Time Training

2025-09-08 · Ruicheng Zhang, Jun Zhou, Zunnan Xu, Zihao Liu 외 arxiv

Trajectory-Guided image-to-video (I2V) generation aims to synthesize videos that adhere to user-specified motion instructions. Existing methods typically rely on computationally expensive fine-tuning on scarce annotated …

Video Generation

Rays as Pixels: Learning A Joint Distribution of Videos and Camera Trajectories

2026-04-10 · Wonbong Jang, Shikun Liu, Soubhik Sanyal, Juan Camilo Perez 외 arxiv

Recovering camera parameters from images and rendering scenes from novel viewpoints have been treated as separate tasks in computer vision and graphics. This separation breaks down when image coverage is sparse or poses …

Video GenerationPose Estimation

FlexTraj: Image-to-Video Generation with Flexible Point Trajectory Control

2025-10-09 · Zhiyuan Zhang, Can Wang, Dongdong Chen, Jing Liao arxiv

We present FlexTraj, a framework for image-to-video generation with flexible point trajectory control. FlexTraj introduces a unified point-based motion representation that encodes each point with a segmentation ID, a tem…

Video Generation

T-GVC: Trajectory-Guided Generative Video Coding at Ultra-Low Bitrates

2025-07-10 · Zhitao Wang, Hengyu Man, Wenrui Li, Xingtao Wang 외 arxiv

Recent advances in video generation techniques have given rise to an emerging paradigm of generative video coding for Ultra-Low Bitrate (ULB) scenarios by leveraging powerful generative priors. However, most existing met…

Video Generation