paper-with-me

Papers

SV3D: Novel Multi-view Synthesis and 3D Generation from a Single Image using Latent Video Diffusion

2024-03-18 · Vikram Voleti, Chun-Han Yao, Mark Boss, Adam Letts, David Pankratz, Dmitry Tochilkin, Christian Laforte, Robin Rombach, Varun Jampani

We present Stable Video 3D (SV3D) -- a latent video diffusion model for high-resolution, image-to-multi-view generation of orbital videos around a 3D object. Recent work on 3D generation propose techniques to adapt 2D generative models for novel view synthesis (NVS) and 3D optimization. However, these methods have several disadvantages due to either limited views or inconsistent NVS, thereby affecting the performance of 3D object generation. In this work, we propose SV3D that adapts image-to-video diffusion model for novel multi-view synthesis and 3D generation, thereby leveraging the generalization and multi-view consistency of the video models, while further adding explicit camera control for NVS. We also propose improved 3D optimization techniques to use SV3D and its NVS outputs for image-to-3D generation. Extensive experimental results on multiple datasets with 2D and 3D metrics as well as user study demonstrate SV3D's state-of-the-art performance on NVS as well as 3D reconstruction compared to prior works.

📄 PDF Abstract BibTeX arXiv:2403.12008

Code (1)

chenguolin/DiffSplat pytorch

Tasks

3D Generation3D ReconstructionImage to 3DNovel View Synthesis

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Exo2EgoSyn: Unlocking Foundation Video Generation Models for Exocentric-to-Egocentric Video Synthesis

2025-11-25 · Mohammad Mahdi, Yuqian Fu, Nedko Savov, Jiancheng Pan 외 arxiv

Foundation video generation models such as WAN 2.2 exhibit strong text- and image-conditioned synthesis abilities but remain constrained to the same-view generation setting. In this work, we introduce Exo2EgoSyn, an adap…

Video Generation

Human Multi-View Synthesis from a Single-View Model:Transferred Body and Face Representations

2024-12-04 · Yu Feng, Shunsi Zhang, Jian Shu, HanFeng Zhao 외

Generating multi-view human images from a single view is a complex and significant challenge. Although recent advancements in multi-view object generation have shown impressive results with diffusion models, novel view s…

Novel View Synthesis

InfiniteNature-Zero: Learning Perpetual View Generation of Natural Scenes from Single Images

2022-07-22 · Zhengqi Li, Qianqian Wang, Noah Snavely, Angjoo Kanazawa

We present a method for learning to generate unbounded flythrough videos of natural scenes starting from a single view, where this capability is learned from a collection of single photographs, without requiring camera p…

Perpetual View Generation

SyncDreamer: Generating Multiview-consistent Images from a Single-view Image

2023-09-07 · YuAn Liu, Cheng Lin, Zijiao Zeng, Xiaoxiao Long 외

In this paper, we present a novel diffusion model called that generates multiview-consistent images from a single-view image. Using pretrained large-scale 2D diffusion models, recent work Zero123 demonstrates the ability…

3D GenerationImage to 3DNovel View SynthesisSingle-View 3D Reconstruction+1

Rotate Your Character: Revisiting Video Diffusion Models for High-Quality 3D Character Generation

2026-01-09 · Jin Wang, Jianxiang Lu, Comi Chen, Guangzheng Xu 외 arxiv

Generating high-quality 3D characters from single images remains a significant challenge in digital content creation, particularly due to complex body poses and self-occlusion. In this paper, we present RCM (Rotate your …

Novel View SynthesisVideo Generation3D Generation