SparseGNV: Generating Novel Views of Indoor Scenes with Sparse Input Views
We study to generate novel views of indoor scenes given sparse input views. The challenge is to achieve both photorealism and view consistency. We present SparseGNV: a learning framework that incorporates 3D structures and image generative models to generate novel views with three modules. The first module builds a neural point cloud as underlying geometry, providing contextual information and guidance for the target novel view. The second module utilizes a transformer-based network to map the scene context and the guidance into a shared latent space and autoregressively decodes the target view in the form of discrete image tokens. The third module reconstructs the tokens into the image of the target view. SparseGNV is trained across a large indoor scene dataset to learn generalizable priors. Once trained, it can efficiently generate novel views of an unseen indoor scene in a feed-forward manner. We evaluate SparseGNV on both real-world and synthetic indoor scenes and demonstrate that it outperforms state-of-the-art methods based on either neural radiance fields or conditional image generation.
Code (1)
Tasks
Conditional Image GenerationImage GenerationSimilar Papers 제목 키워드 기반
Planar Surface Reconstruction from Sparse Views
The paper studies planar surface reconstruction of indoor scenes from two views with unknown camera poses. While prior approaches have successfully created object-centric reconstructions of many scenes, they fail to expl…
Surface ReconstructionGlobal and Hierarchical Geometry Consistency Priors for Few-shot NeRFs in Indoor Scenes
It is challenging for Neural Radiance Fields (NeRFs) in the few-shot setting to reconstruct high-quality novel views and depth maps in 360^\circ outward-facing indoor scenes. The captured sparse views for these scene…
Depth EstimationDepth PredictionMonocular Depth EstimationNeRFLearning to Reconstruct and Understand Indoor Scenes from Sparse Views
This paper proposes a new method for simultaneous 3D reconstruction and semantic segmentation of indoor scenes. Unlike existing methods that require recording a video using a color camera and/or a depth camera, our metho…
3D ReconstructionDepth EstimationSegmentationSemantic SegmentationSparis: Neural Implicit Surface Reconstruction of Indoor Scenes from Sparse Views
In recent years, reconstructing indoor scene geometry from multi-view images has achieved encouraging accomplishments. Current methods incorporate monocular priors into neural implicit surface models to achieve high-qual…
Surface ReconstructionDense Depth Priors for Neural Radiance Fields from Sparse Input Views
Neural radiance fields (NeRF) encode a scene into a neural representation that enables photo-realistic rendering of novel views. However, a successful reconstruction from RGB images requires a large number of input views…
Depth CompletionNeRFNovel View Synthesis