paper-with-me

홈 › Papers

SPATIALGEN: Layout-guided 3D Indoor Scene Generation

2025-09-18 · Chuan Fang, Heng Li, Yixun Liang, Jia Zheng, Yongsen Mao, Yuan Liu, Rui Tang, Zihan Zhou, Ping Tan arxiv

Creating high-fidelity 3D models of indoor environments is essential for applications in design, virtual reality, and robotics. However, manual 3D modeling remains time-consuming and labor-intensive. While recent advances in generative AI have enabled automated scene synthesis, existing methods often face challenges in balancing visual quality, diversity, semantic consistency, and user control. A major bottleneck is the lack of a large-scale, high-quality dataset tailored to this task. To address this gap, we introduce a comprehensive synthetic dataset, featuring 12,328 structured annotated scenes with 57,431 rooms, and 4.7M photorealistic 2D renderings. Leveraging this dataset, we present SpatialGen, a novel multi-view multi-modal diffusion model that generates realistic and semantically consistent 3D indoor scenes. Given a 3D layout and a reference image (derived from a text prompt), our model synthesizes appearance (color image), geometry (scene coordinate map), and semantic (semantic segmentation map) from arbitrary viewpoints, while preserving spatial consistency across modalities. SpatialGen consistently generates superior results to previous methods in our experiments. We are open-sourcing our data and models to empower the community and advance the field of indoor scene understanding and generation.

📄 PDF Abstract BibTeX arXiv:2509.14981

Code (0)

등록된 구현이 없습니다.

Tasks

Semantic SegmentationScene UnderstandingScene Generation

Similar Papers 제목 키워드 기반

Scene Generation at Absolute Scale: Utilizing Semantic and Geometric Guidance From Text for Accurate and Interpretable 3D Indoor Scene Generation

2026-03-14 · Stefan Ainetter, Thomas Deixelberger, Edoardo A. Dominici, Philipp Drescher 외 arxiv

We present GuidedSceneGen, a text-to-3D generation framework that produces metrically accurate, globally consistent, and semantically interpretable indoor scenes. Unlike prior text-driven methods that often suffer from g…

Collision AvoidanceScene Generation3D Generation

SceneCraft: Layout-Guided 3D Scene Generation

2024-10-11 · Xiuyu Yang, Yunze Man, Jun-Kun Chen, Yu-Xiong Wang

The creation of complex 3D scenes tailored to user specifications has been a tedious and challenging task with traditional 3D modeling tools. Although some pioneering methods have achieved automatic text-to-3D generation…

3D GenerationImage GenerationNeRFScene Generation+1

HetScene: Heterogeneity-Aware Diffusion for Dense Indoor Scene Generation

2026-05-13 · Zini Chen, Junming Huang, Rong Zhang, Jiamin Xu 외 arxiv

Generating controllable and physically plausible indoor scenes is a pivotal prerequisite for constructing high-fidelity simulation environments for embodied AI. However, existing deeplearning-based methods usually treat …

Scene Generation

Ctrl-Room: Controllable Text-to-3D Room Meshes Generation with Layout Constraints

2023-10-05 · Chuan Fang, Yuan Dong, Kunming Luo, Xiaotao Hu 외

Text-driven 3D indoor scene generation is useful for gaming, the film industry, and AR/VR applications. However, existing methods cannot faithfully capture the room layout, nor do they allow flexible editing of individua…

Layout GenerationScene GenerationText to 3D

InSpace: Structure-Aware 3D Indoor Scene Generation from a Single 360° Image

2026-07-04 · Gwanhyeong Koo, Hyunsu Kim, Youngji Kim, Taejae Lee 외 arxiv

Recent advances in single image-to-3D generation have enabled high-quality asset synthesis, yet extending these capabilities to indoor scene generation remains challenging. Existing methods focus on asset-level generatio…

Scene Generation3D Generation