paper-with-me

Papers

SynCity 3000: Bootstrapping Scene-Scale 3D Diffusion

2026-07-06 · Paul Engstler, Iro Laina, Christian Rupprecht, Andrea Vedaldi arxiv

We present SynCity 3000, a framework for generating 3D scenes that are globally coherent while enabling fine-grained layout control. Building on the ability of current image-to-3D generators to produce complex 3D assets from a single image, we extend this capability to the scale of entire scenes by adapting the generator to be applicable as a convolutional operator. We achieve this by fine-tuning the model on scene-like data generated by a new synthetic data engine, which we propose to address the scarcity of 3D scene data for training. The convolutional generator is then applied to a dimetric image of the entire scene, generated from the user prompt, resulting in 3D scenes of arbitrary size and complexity. Across diverse prompts and layouts, SynCity 3000 produces large, coherent, and detailed scenes, addressing the shortcomings of prior approaches to 3D scene generation.

📄 PDF Abstract BibTeX arXiv:2607.05392

Code (0)

등록된 구현이 없습니다.

Tasks

Scene Generation

Similar Papers 제목 키워드 기반

SynCity: Training-Free Generation of 3D Worlds

2025-03-20 · Paul Engstler, Aleksandar Shtedritski, Iro Laina, Christian Rupprecht 외

We address the challenge of generating 3D worlds from textual descriptions. We propose SynCity, a training- and optimization-free approach, which leverages the geometric precision of pre-trained 3D generative models and …

Diversity

ObjectDrop: Bootstrapping Counterfactuals for Photorealistic Object Removal and Insertion

2024-03-27 · Daniel Winter, Matan Cohen, Shlomi Fruchter, Yael Pritch 외

Diffusion models have revolutionized image editing but often generate images that violate physical laws, particularly the effects of objects on the scene, e.g., occlusions, shadows, and reflections. By analyzing the limi…

counterfactualObject

Diffusion Probabilistic Models for Scene-Scale 3D Categorical Data

2023-01-02 · Jumin Lee, Woobin Im, Sebin Lee, Sung-Eui Yoon

In this paper, we learn a diffusion model to generate 3D data on a scene-scale. Specifically, our model crafts a 3D scene consisting of multiple objects, while recent diffusion research has focused on a single object. To…

Accurate Scene Text Detection through Border Semantics Awareness and Bootstrapping

2018-07-10 · ECCV 2018 9 · Chuhui Xue, Shijian Lu, Fangneng Zhan

This paper presents a scene text detection technique that exploits bootstrapping and text border semantics for accurate localization of texts in scenes. A novel bootstrapping technique is designed which samples multiple …

Scene Text DetectionText Detection

BOOT: Data-free Distillation of Denoising Diffusion Models with Bootstrapping

2023-06-08 · Jiatao Gu, Shuangfei Zhai, Yizhe Zhang, Lingjie Liu 외

Diffusion models have demonstrated excellent potential for generating diverse images. However, their performance often suffers from slow generation due to iterative denoising. Knowledge distillation has been recently pro…

DenoisingKnowledge Distillation