paper-with-me

Papers

One2Scene: Geometric Consistent Explorable 3D Scene Generation from a Single Image

2026-02-23 · Pengfei Wang, Liyi Chen, Zhiyuan Ma, Yanjun Guo, Guowen Zhang, Lei Zhang arxiv

Generating explorable 3D scenes from a single image is a highly challenging problem in 3D vision. Existing methods struggle to support free exploration, often producing severe geometric distortions and noisy artifacts when the viewpoint moves far from the original perspective. We introduce \textbf{One2Scene}, an effective framework that decomposes this ill-posed problem into three tractable sub-tasks to enable immersive explorable scene generation. We first use a panorama generator to produce anchor views from a single input image as initialization. Then, we lift these 2D anchors into an explicit 3D geometric scaffold via a generalizable, feed-forward Gaussian Splatting network. Instead of treating the panorama as a single image for reconstruction, we project it into multiple sparse anchor views and reformulate the reconstruction task as multi-view stereo matching, which allows us to leverage robust geometric priors learned from large-scale multi-view datasets. A bidirectional feature fusion module is used to enforce cross-view consistency, yielding an efficient and geometrically reliable scaffold. Finally, the scaffold serves as a strong prior for a novel view generator to produce photorealistic and geometrically accurate views at arbitrary cameras. By explicitly conditioning on a 3D-consistent scaffold to perform reconstruction, One2Scene works stably under large camera motions, supporting immersive scene exploration. Extensive experiments show that One2Scene substantially outperforms state-of-the-art methods in panorama depth estimation, feed-forward 360° reconstruction, and explorable 3D scene generation. Project page: https://one2scene5406.github.io/

📄 PDF Abstract BibTeX arXiv:2602.19766

Code (0)

등록된 구현이 없습니다.

Tasks

Depth EstimationScene Generation

Similar Papers 제목 키워드 기반

Matrix-3D: Omnidirectional Explorable 3D World Generation

2025-08-11 · Zhongqi Yang, Wenhang Ge, Yuqi Li, Jiaqi Chen 외 arxiv

Explorable 3D world generation from a single image or text prompt forms a cornerstone of spatial intelligence. Recent works utilize video model to achieve wide-scope and generalizable 3D world generation. However, existi…

3D ReconstructionVideo Generation

Voyager: Long-Range and World-Consistent Video Diffusion for Explorable 3D Scene Generation

2025-06-04 · Tianyu Huang, Wangguandong Zheng, Tengfei Wang, Yuhao Liu 외

Real-world applications like video gaming and virtual reality often demand the ability to model 3D scenes that users can explore along custom camera trajectories. While significant progress has been made in generating 3D…

3D ReconstructionCamera Pose EstimationDepth EstimationDepth Prediction+3

SpatialCrafter: Single Image World Modeling with Generative 3D Proxies

2026-08-27 · Chuan Fang, Lingteng Qiu, Yixun Liang, Rui Chen 외 arxiv

Explorable image-to-scene generation is essential for applications in gaming, robotics, and virtual reality. Existing methods based on video diffusion model (VDM) commonly rely on incomplete conditioning signals such as …

Scene GenerationPoint Clouds

Lyra 2.0: Explorable Generative 3D Worlds

2026-04-14 · Tianchang Shen, Sherwin Bahmani, Kai He, Sangeetha Grama Srinivasan 외 arxiv

Recent advances in video generation enable a new paradigm for 3D scene creation: generating camera-controlled videos that simulate scene walkthroughs, then lifting them to 3D via feed-forward reconstruction techniques. T…

Video Generation

Skyfall-GS: Synthesizing Immersive 3D Urban Scenes from Satellite Imagery

2025-10-17 · Jie-Ying Lee, Yi-Ruei Liu, Shr-Ruei Tsai, Wei-Cheng Chang 외 arxiv

Synthesizing large-scale, explorable, and geometrically accurate 3D urban scenes is a challenging yet valuable task for immersive and embodied applications. The challenge lies in the lack of large-scale and high-quality …