paper-with-me

홈 › Papers

DiffInDScene: Diffusion-based High-Quality 3D Indoor Scene Generation

2023-06-01 · CVPR 2024 1 · Xiaoliang Ju, Zhaoyang Huang, Yijin Li, Guofeng Zhang, Yu Qiao, Hongsheng Li

We present DiffInDScene, a novel framework for tackling the problem of high-quality 3D indoor scene generation, which is challenging due to the complexity and diversity of the indoor scene geometry. Although diffusion-based generative models have previously demonstrated impressive performance in image generation and object-level 3D generation, they have not yet been applied to room-level 3D generation due to their computationally intensive costs. In DiffInDScene, we propose a cascaded 3D diffusion pipeline that is efficient and possesses strong generative performance for Truncated Signed Distance Function (TSDF). The whole pipeline is designed to run on a sparse occupancy space in a coarse-to-fine fashion. Inspired by KinectFusion's incremental alignment and fusion of local TSDF volumes, we propose a diffusion-based SDF fusion approach that iteratively diffuses and fuses local TSDF volumes, facilitating the generation of an entire room environment. The generated results demonstrate that our work is capable to achieve high-quality room generation directly in three-dimensional space, starting from scratch. In addition to the scene generation, the final part of DiffInDScene can be used as a post-processing module to refine the 3D reconstruction results from multi-view stereo. According to the user study, the mesh quality generated by our DiffInDScene can even outperform the ground truth mesh provided by ScanNet. Please visit our project page for the latest progress and demonstrations: https://github.com/AkiraHero/diffindscene.

📄 PDF Abstract BibTeX arXiv:2306.00519

Code (1)

akirahero/diffindscene 공식 구현 pytorch

Tasks

3D Generation3D ReconstructionImage GenerationScene Generation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

SceneTex: High-Quality Texture Synthesis for Indoor Scenes via Diffusion Priors

2023-11-28 · CVPR 2024 1 · Dave Zhenyu Chen, Haoxuan Li, Hsin-Ying Lee, Sergey Tulyakov 외

We propose SceneTex, a novel method for effectively generating high-quality and style-consistent textures for indoor scenes using depth-to-image diffusion priors. Unlike previous methods that either iteratively warp 2D v…

DecoderTexture Synthesis

Mixed Diffusion for 3D Indoor Scene Synthesis

2024-05-31 · Siyi Hu, Diego Martin Arroyo, Stephanie Debats, Fabian Manhardt 외

Generating realistic 3D scenes is an area of growing interest in computer vision and robotics. However, creating high-quality, diverse synthetic 3D content often requires expert intervention, making it costly and complex…

DenoisingIndoor Scene Synthesis

InsTex: Indoor Scenes Stylized Texture Synthesis

2025-01-22 · Yunfan Zhang, Zhiwei Xiong, Zhiqi Shen, Guosheng Lin 외

Generating high-quality textures for 3D scenes is crucial for applications in interior design, gaming, and augmented/virtual reality (AR/VR). Although recent advancements in 3D generative models have enhanced content cre…

Texture Synthesis

CommonScenes: Generating Commonsense 3D Indoor Scenes with Scene Graph Diffusion

2023-05-25 · NeurIPS 2023 11 · Guangyao Zhai, Evin Pınar Örnek, Shun-Cheng Wu, Yan Di 외

Controllable scene synthesis aims to create interactive environments for various industrial use cases. Scene graphs provide a highly suitable interface to facilitate these applications by abstracting the scene context in…

DiversityObjectScene Generation

Rein3D: Reinforced 3D Indoor Scene Generation with Panoramic Video Diffusion Models

2026-04-12 · Dehui Wang, Rong Wei, Yue Shi, Congsheng Xu 외 arxiv

The growing demand for Embodied AI and VR applications has highlighted the need for synthesizing high-quality 3D indoor scenes from sparse inputs. However, existing approaches struggle to infer massive amounts of missing…

Video Super-ResolutionScene Generation