paper-with-me

Papers

InstructLayout: Instruction-Driven 2D and 3D Layout Synthesis with Semantic Graph Prior

2024-07-10 · Chenguo Lin, YuChen Lin, Panwang Pan, Xuanyang Zhang, Yadong Mu

Comprehending natural language instructions is a charming property for both 2D and 3D layout synthesis systems. Existing methods implicitly model object joint distributions and express object relations, hindering generation's controllability. We introduce InstructLayout, a novel generative framework that integrates a semantic graph prior and a layout decoder to improve controllability and fidelity for 2D and 3D layout synthesis. The proposed semantic graph prior learns layout appearances and object distributions simultaneously, demonstrating versatility across various downstream tasks in a zero-shot manner. To facilitate the benchmarking for text-driven 2D and 3D scene synthesis, we respectively curate two high-quality datasets of layout-instruction pairs from public Internet resources with large language and multimodal models. Extensive experimental results reveal that the proposed method outperforms existing state-of-the-art approaches by a large margin in both 2D and 3D layout synthesis tasks. Thorough ablation studies confirm the efficacy of crucial design components.

📄 PDF Abstract BibTeX arXiv:2407.07580

Code (1)

chenguolin/InstructScene pytorch

Tasks

BenchmarkingDecoderObject

Similar Papers 제목 키워드 기반

InstructScene: Instruction-Driven 3D Indoor Scene Synthesis with Semantic Graph Prior

2024-02-07 · Chenguo Lin, Yadong Mu

Comprehending natural language instructions is a charming property for 3D indoor scene synthesis systems. Existing methods directly model object joint distributions and express object relations implicitly within a scene,…

BenchmarkingDecoderIndoor Scene Synthesis

LoCo: Locally Constrained Training-Free Layout-to-Image Synthesis

2023-11-21 · Peiang Zhao, Han Li, Ruiyang Jin, S. Kevin Zhou

Recent text-to-image diffusion models have reached an unprecedented level in generating high-quality images. However, their exclusive reliance on textual prompts often falls short in precise control of image compositions…

Image Generation

ReSpace: Text-Driven 3D Scene Synthesis and Editing with Preference Alignment

2025-06-03 · Martin JJ. Bucher, Iro Armeni

Scene synthesis and editing has emerged as a promising direction in computer graphics. Current trained approaches for 3D indoor scenes either oversimplify object semantics through one-hot class encodings (e.g., 'chair' o…

Indoor Scene SynthesisObjectSpatial Reasoning

M3DLayout: A Multi-Source Dataset of 3D Indoor Layouts and Structured Descriptions for 3D Generation

2025-09-28 · Yiheng Zhang, Zhuojiang Cai, Mingdao Wang, Meitong Guo 외 arxiv

In text-driven 3D scene generation, object layout serves as a crucial intermediate representation that bridges high-level language instructions with detailed geometric output. It not only provides a structural blueprint …

Scene Generation3D Generation

Controllable Generation of Large-Scale 3D Urban Layouts with Semantic and Structural Guidance

2025-09-28 · Mengyuan Niu, Xinxin Zhuo, Ruizhe Wang, Yuyue Huang 외 arxiv

Urban modeling is essential for city planning, scene synthesis, and gaming. Existing image-based methods generate diverse layouts but often lack geometric continuity and scalability, while graph-based methods capture str…