paper-with-me

홈 › Papers

CameraNoise: Enabling Faithful Camera Control in Video Diffusion through Geometry-Flow-Guided Noise Warping

2026-05-29 · Haoyu Zhao, Jiaxi Gu, Haoran Chen, Qingping Zheng, Yeying Jin, Hongyi Yang, Junqi Cheng, Yuang Zhang, Zenghui Lu, Huan Yu, Jie Jiang, Peng Shu, Zuxuan Wu, Yu-Gang Jiang arxiv

Precise camera pose control is critical for video diffusion, yet maintaining geometric consistency remains a challenge. Existing methods that directly inject numerical camera parameters into the diffusion backbone often fail to bridge the gap between abstract coordinates and visual content, leading to structural distortions. To address this issue, we propose CameraNoise, a flow-to-noise warping method that encodes camera motion into a temporally coherent stochastic representation. Unlike conventional conditioning, CameraNoise embeds camera poses directly into the noise space. This decouples motion from scene appearance while faithfully preserving trajectory dynamics. Specifically, we introduce a novel Geometry-guided Reprojection Flow and a noise warping algorithm, which jointly preserve the Gaussian prior of diffusion and ensure consistent noise propagation under camera transformations. By integrating CameraNoise into the diffusion process, our framework delivers stable, high-fidelity videos. Extensive experiments demonstrate that our approach significantly outperforms prior methods in both visual quality and trajectory faithfulness. The project page and code are available at: https://gulucaptain.github.io/CameraNoise/.

📄 PDF Abstract BibTeX arXiv:2605.30774

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Beyond Inpainting: Unleash 3D Understanding for Precise Camera-Controlled Video Generation

2026-01-15 · Dong-Yu Chen, Yixin Guo, Shuojin Yang, Tai-Jiang Mu 외 arxiv

Camera control has been extensively studied in conditioned video generation; however, performing precisely altering the camera trajectories while faithfully preserving the video content remains a challenging task. The ma…

Video Generation

VerseCrafter: Dynamic Realistic Video World Model with 4D Geometric Control

2026-01-08 · Sixiao Zheng, Minghao Yin, Wenbo Hu, Xiaoyu Li 외 arxiv

Video world models aim to simulate dynamic, real-world environments, yet existing methods struggle to provide unified and precise control over camera and multi-object motion, as videos inherently capture dynamics in the …

Video Generation

CT-1: Vision-Language-Camera Models Transfer Spatial Reasoning Knowledge to Camera-Controllable Video Generation

2026-04-10 · Haoyu Zhao, Zihao Zhang, Jiaxi Gu, Haoran Chen 외 arxiv

Camera-controllable video generation aims to synthesize videos with flexible and physically plausible camera movements. However, existing methods either provide imprecise camera control from text prompts or rely on labor…

Spatial ReasoningVideo Generation

CameraCtrl: Enabling Camera Control for Text-to-Video Generation

2024-04-02 · Hao He, Yinghao Xu, Yuwei Guo, Gordon Wetzstein 외

Controllability plays a crucial role in video generation since it allows users to create desired content. However, existing models largely overlooked the precise control of camera pose that serves as a cinematic language…

Text-to-Video GenerationVideo Generation

Infinite-Homography as Robust Conditioning for Camera-Controlled Video Generation

2025-12-18 · Min-Jung Kim, Jeongho Kim, Hoiyeong Jin, Junha Hyung 외 arxiv

Recent progress in video diffusion models has spurred growing interest in camera-controlled novel-view video generation for dynamic scenes, aiming to provide creators with cinematic camera control capabilities in post-pr…

Data AugmentationDepth EstimationVideo Generation