paper-with-me

Papers

Collaborative Video Diffusion: Consistent Multi-video Generation with Camera Control

2024-05-27 · Zhengfei Kuang, Shengqu Cai, Hao He, Yinghao Xu, Hongsheng Li, Leonidas Guibas, Gordon Wetzstein

Research on video generation has recently made tremendous progress, enabling high-quality videos to be generated from text prompts or images. Adding control to the video generation process is an important goal moving forward and recent approaches that condition video generation models on camera trajectories make strides towards it. Yet, it remains challenging to generate a video of the same scene from multiple different camera trajectories. Solutions to this multi-video generation problem could enable large-scale 3D scene generation with editable camera trajectories, among other applications. We introduce collaborative video diffusion (CVD) as an important step towards this vision. The CVD framework includes a novel cross-video synchronization module that promotes consistency between corresponding frames of the same video rendered from different camera poses using an epipolar attention mechanism. Trained on top of a state-of-the-art camera-control module for video generation, CVD generates multiple videos rendered from different camera trajectories with significantly better consistency than baselines, as shown in extensive experiments. Project page: https://collaborativevideodiffusion.github.io/.

📄 PDF Abstract BibTeX arXiv:2405.17414

Code (0)

등록된 구현이 없습니다.

Tasks

Scene GenerationVideo GenerationVideo Synchronization

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

ANYPORTAL: Zero-Shot Consistent Video Background Replacement

2025-09-09 · Wenshuo Gao, Xicheng Lan, Shuai Yang arxiv

Despite the rapid advancements in video generation technology, creating high-quality videos that precisely align with user intentions remains a significant challenge. Existing methods often fail to achieve fine-grained c…

Video Generation

Collaborative Score Distillation for Consistent Visual Synthesis

2023-07-04 · Subin Kim, Kyungmin Lee, June Suk Choi, Jongheon Jeong 외

Generative priors of large-scale text-to-image diffusion models enable a wide range of new generation and editing applications on diverse visual modalities. However, when adapting these priors to complex visual modalitie…

Collaborative Score Distillation for Consistent Visual Editing

2023-09-21 · NeurIPS 2023 11

Generative priors of large-scale text-to-image diffusion models enable a wide range of new generation and editing applications on diverse visual modalities. However, when adapting these priors to complex visual modalitie…

View-Consistent Diffusion Representations for 3D-Consistent Video Generation

2025-11-24 · Duolikun Danier, Ge Gao, Steven McDonagh, Changjian Li 외 arxiv

Video generation models have made significant progress in generating realistic content, enabling applications in simulation, gaming, and film making. However, current generated videos still contain visual artifacts arisi…

Video Generation

Tuning-Free Long Video Generation via Global-Local Collaborative Diffusion

2025-01-08 · Yongjia Ma, Junlin Chen, Donglin Di, Qi Xie 외

Creating high-fidelity, coherent long videos is a sought-after aspiration. While recent video diffusion models have shown promising potential, they still grapple with spatiotemporal inconsistencies and high computational…

DenoisingDiversityVideo DenoisingVideo Generation