paper-with-me

홈 › Papers

SceneGen: Single-Image 3D Scene Generation in One Feedforward Pass

2025-08-21 · Yanxu Meng, Haoning Wu, Ya Zhang, Weidi Xie arxiv

3D content generation has recently attracted significant research interest, driven by its critical applications in VR/AR and embodied AI. In this work, we tackle the challenging task of synthesizing multiple 3D assets within a single scene image. Concretely, our contributions are fourfold: (i) we present SceneGen, a novel framework that takes a scene image and corresponding object masks as input, simultaneously producing multiple 3D assets with geometry and texture. Notably, SceneGen operates with no need for extra optimization or asset retrieval; (ii) we introduce a novel feature aggregation module that integrates local and global scene information from visual and geometric encoders within the feature extraction module. Coupled with a position head, this enables the generation of 3D assets and their relative spatial positions in a single feedforward pass; (iii) we demonstrate SceneGen's direct extensibility to multi-image input scenarios. Despite being trained solely on single-image inputs, our architecture yields improved generation performance when multiple images are provided; and (iv) extensive quantitative and qualitative evaluations confirm the efficiency and robustness of our approach. We believe this paradigm offers a novel solution for high-quality 3D content generation, potentially advancing its practical applications in downstream tasks. The code and model will be publicly available at: https://mengmouxu.github.io/SceneGen.

📄 PDF Abstract BibTeX arXiv:2508.15769

Code (0)

등록된 구현이 없습니다.

Tasks

Scene Generation

Similar Papers 제목 키워드 기반

SceneGenAgent: Precise Industrial Scene Generation with Coding Agent

2024-10-29 · Xiao Xia, Dan Zhang, Zibo Liao, Zhenyu Hou 외

The modeling of industrial scenes is essential for simulations in industrial manufacturing. While large language models (LLMs) have shown significant progress in generating general 3D scenes from textual descriptions, ge…

C++ codeScene Generation

DriveSceneGen: Generating Diverse and Realistic Driving Scenarios from Scratch

2023-09-26 · Shuo Sun, Zekai Gu, Tianchen Sun, Jiawei Sun 외

Realistic and diverse traffic scenarios in large quantities are crucial for the development and validation of autonomous driving systems. However, owing to numerous difficulties in the data collection process and the rel…

Autonomous DrivingDiversity

SceneGen: Learning to Generate Realistic Traffic Scenes

2021-01-16 · CVPR 2021 1 · Shuhan Tan, Kelvin Wong, Shenlong Wang, Sivabalan Manivasagam 외

We consider the problem of generating realistic traffic scenes automatically. Existing methods typically insert actors into the scene according to a set of hand-crafted heuristics and are limited in their ability to mode…

Diversity

Scene Generation at Absolute Scale: Utilizing Semantic and Geometric Guidance From Text for Accurate and Interpretable 3D Indoor Scene Generation

2026-03-14 · Stefan Ainetter, Thomas Deixelberger, Edoardo A. Dominici, Philipp Drescher 외 arxiv

We present GuidedSceneGen, a text-to-3D generation framework that produces metrically accurate, globally consistent, and semantically interpretable indoor scenes. Unlike prior text-driven methods that often suffer from g…

Collision AvoidanceScene Generation3D Generation

PosterMaker: Towards High-Quality Product Poster Generation with Accurate Text Rendering

2025-04-09 · CVPR 2025 1 · Yifan Gao, Zihang Lin, Chuanbin Liu, Min Zhou 외

Product posters, which integrate subject, scene, and text, are crucial promotional tools for attracting customers. Creating such posters using modern image generation methods is valuable, while the main challenge lies in…

Image Generation