paper-with-me

홈 › Papers

WorldGrow: Generating Infinite 3D World

2025-10-24 · Sikuang Li, Chen Yang, Jiemin Fang, Taoran Yi, Jia Lu, Jiazhong Cen, Lingxi Xie, Wei Shen, Qi Tian arxiv

We tackle the challenge of generating the infinitely extendable 3D world -- large, continuous environments with coherent geometry and realistic appearance. Existing methods face key challenges: 2D-lifting approaches suffer from geometric and appearance inconsistencies across views, 3D implicit representations are hard to scale up, and current 3D foundation models are mostly object-centric, limiting their applicability to scene-level generation. Our key insight is leveraging strong generation priors from pre-trained 3D models for structured scene block generation. To this end, we propose WorldGrow, a hierarchical framework for unbounded 3D scene synthesis. Our method features three core components: (1) a data curation pipeline that extracts high-quality scene blocks for training, making the 3D structured latent representations suitable for scene generation; (2) a 3D block inpainting mechanism that enables context-aware scene extension; and (3) a coarse-to-fine generation strategy that ensures both global layout plausibility and local geometric/textural fidelity. Evaluated on the large-scale 3D-FRONT dataset, WorldGrow achieves SOTA performance in geometry reconstruction, while uniquely supporting infinite scene generation with photorealistic and structurally consistent outputs. These results highlight its capability for constructing large-scale virtual environments and potential for building future world models.

📄 PDF Abstract BibTeX arXiv:2510.21682

Code (0)

등록된 구현이 없습니다.

Tasks

Scene Generation

Similar Papers 제목 키워드 기반

Infinite Structured Hidden Semi-Markov Models

2014-06-30 · Jonathan H. Huggins, Frank Wood

This paper reviews recent advances in Bayesian nonparametric techniques for constructing and performing inference in infinite hidden Markov models. We focus on variants of Bayesian nonparametric hidden Markov models that…

InfiniteWeb: Scalable Web Environment Synthesis for GUI Agent Training

2026-01-07 · Ziyun Zhang, Zezhou Wang, Xiaoyi Zhang, Zongyu Guo 외 arxiv

GUI agents that interact with graphical interfaces on behalf of users represent a promising direction for practical AI assistants. However, training such agents is hindered by the scarcity of suitable environments. We pr…

Reinforcement Learning

InfiniteAudio: Infinite-Length Audio Generation with Consistency

2025-06-03 · Chaeyoung Jung, Hojoon Ki, Ji-Hoon Kim, Junmo Kim 외

This paper presents InfiniteAudio, a simple yet effective strategy for generating infinite-length audio using diffusion-based text-to-audio methods. Current approaches face memory constraints because the output size incr…

Audio GenerationDenoising

MagicInfinite: Generating Infinite Talking Videos with Your Words and Voice

2025-03-07 · Hongwei Yi, Tian Ye, Shitong Shao, Xuancheng Yang 외

We present MagicInfinite, a novel diffusion Transformer (DiT) framework that overcomes traditional portrait animation limitations, delivering high-fidelity results across diverse character types-realistic humans, full-bo…

DenoisingPortrait AnimationVideo Generation

Probabilistic Inference with Generating Functions for Poisson Latent Variable Models

2016-12-01 · NeurIPS 2016 12 · Kevin Winner, Daniel R. Sheldon

Graphical models with latent count variables arise in a number of fields. Standard exact inference techniques such as variable elimination and belief propagation do not apply to these models because the latent variables …