paper-with-me

Papers

Stable Virtual Camera: Generative View Synthesis with Diffusion Models

2025-03-18 · Jensen, Zhou, Hang Gao, Vikram Voleti, Aaryaman Vasishta, Chun-Han Yao, Mark Boss, Philip Torr, Christian Rupprecht, Varun Jampani

We present Stable Virtual Camera (Seva), a generalist diffusion model that creates novel views of a scene, given any number of input views and target cameras. Existing works struggle to generate either large viewpoint changes or temporally smooth samples, while relying on specific task configurations. Our approach overcomes these limitations through simple model design, optimized training recipe, and flexible sampling strategy that generalize across view synthesis tasks at test time. As a result, our samples maintain high consistency without requiring additional 3D representation-based distillation, thus streamlining view synthesis in the wild. Furthermore, we show that our method can generate high-quality videos lasting up to half a minute with seamless loop closure. Extensive benchmarking demonstrates that Seva outperforms existing methods across different datasets and settings.

📄 PDF Abstract BibTeX arXiv:2503.14489

Code (0)

등록된 구현이 없습니다.

Tasks

Benchmarking

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

FreeStyleGAN: Free-view Editable Portrait Rendering with the Camera Manifold

2021-09-20 · Thomas Leimkühler, George Drettakis

Current Generative Adversarial Networks (GANs) produce photorealistic renderings of portrait images. Embedding real images into the latent space of such models enables high-level image editing. While recent methods provi…

3D ReconstructionMixed RealityNovel View Synthesis

View Synthesis of Dynamic Scenes based on Deep 3D Mask Volume

2021-08-30 · ICCV 2021 10 · Kai-En Lin, Guowei Yang, Lei Xiao, Feng Liu 외

Image view synthesis has seen great success in reconstructing photorealistic visuals, thanks to deep learning and various novel representations. The next key step in immersive virtual experiences is view synthesis of dyn…

InfiniteNature-Zero: Learning Perpetual View Generation of Natural Scenes from Single Images

2022-07-22 · Zhengqi Li, Qianqian Wang, Noah Snavely, Angjoo Kanazawa

We present a method for learning to generate unbounded flythrough videos of natural scenes starting from a single view, where this capability is learned from a collection of single photographs, without requiring camera p…

Perpetual View Generation

Generative Camera Dolly: Extreme Monocular Dynamic Novel View Synthesis

2024-05-23 · Basile Van Hoorick, Rundi Wu, Ege Ozguroglu, Kyle Sargent 외

Accurate reconstruction of complex dynamic scenes from just a single viewpoint continues to be a challenging task in computer vision. Current dynamic novel view synthesis methods typically require videos from many differ…

Novel View SynthesisScene Understanding

Non-Rigid Neural Radiance Fields: Reconstruction and Novel View Synthesis of a Dynamic Scene From Monocular Video

2020-12-22 · ICCV 2021 10 · Edgar Tretschk, Ayush Tewari, Vladislav Golyanik, Michael Zollhöfer 외

We present Non-Rigid Neural Radiance Fields (NR-NeRF), a reconstruction and novel view synthesis approach for general non-rigid dynamic scenes. Our approach takes RGB images of a dynamic scene as input (e.g., from a mono…

NeRFNovel View SynthesisVideo Editing