paper-with-me

홈 › Papers

Full-4D: Generating Full-Scope 4D Scenes from a Single-View Video

2026-05-25 · Tingxi Chen, Ke Hao, Yabo Chen, Zhengxue Cheng, Rong Xie, Li Song, Haibin Huang, Chi Zhang, Xuelong Li arxiv

Generating 4D scenes from a single-view video is inherently ill-posed: a single viewpoint lacks the information needed to recover a complete, dynamic scene with full coverage. Existing methods are typically limited to monocular videos, simple 3D effects, or only small viewpoint perturbations around the original viewpoint, falling short of true 4D generation. Meanwhile, the lack of large-scale datasets capturing full-scope 4D scenes with synchronized multi-view videos further hinders progress in this direction. We propose a novel single-view video-to-4D framework that casts full-scope 4D generation as a multi-view video synthesis followed by optimization-based 4D reconstruction from the generated views. To instantiate this formulation end-to-end, we make three key contributions. First, we introduce Real-MV-4D, a large-scale dataset of synchronized multi-view videos captured in diverse real-world environments to provide the 4D supervision. Second, we train a multi-view video diffusion model driven by a novel fused time(T)-view(V) attention mechanism that directly embeds geometric reprojection priors and explicit camera conditioning into its view-time interactions. Unlike basic feature fusion, this direct binding strictly aligns the generation process with physical 3D priors to produce a dense, synchronized T$\times $V video grid. Third, rather than relying on non-interactive and inconsistent 2D video interpolations, we lift the synthesized multi-view videos into an explicit 4D representation (i.e. 4DGS), regularized by a Flow Matching Distillation loss that exploits the multi-view prior to improve novel-view rendering. Extensive experiments demonstrate that our method outperforms existing approaches in both visual fidelity and geometric consistency, enabling full-scope 4D scene generation from single-view videos.

📄 PDF Abstract BibTeX arXiv:2605.25500

Code (0)

등록된 구현이 없습니다.

Tasks

Scene Generation

Similar Papers 제목 키워드 기반

GEN3D: Generating Domain-Free 3D Scenes from a Single Image

2025-11-18 · Yuxin Zhang, Ziyu Lu, Hongbo Duan, Keyu Fan 외 arxiv

Despite recent advancements in neural 3D reconstruction, the dependence on dense multi-view captures restricts their broader applicability. Additionally, 3D scene generation is vital for advancing embodied AI and world m…

3D ReconstructionScene Generation

Neural Kaleidoscopic Space Sculpting

2023-01-01 · CVPR 2023 1 · Byeongjoo Ahn, Michael De Zeeuw, Ioannis Gkioulekas, Aswin C. Sankaranarayanan

We introduce a method that recovers full-surround 3D reconstructions from a single kaleidoscopic image using a neural surface representation. Full-surround 3D reconstruction is critical for many applications, such as…

3D Reconstruction

ARShadowGAN: Shadow Generative Adversarial Network for Augmented Reality in Single Light Scenes

2020-06-01 · CVPR 2020 6 · Daquan Liu, Chengjiang Long, Hongpan Zhang, Hanning Yu 외

Generating virtual object shadows consistent with the real-world environment shading effects is important but challenging in computer vision and augmented reality applications. To address this problem, we propose an end-…

Generative Adversarial NetworkObject

MOGRAS: Human Motion with Grasping in 3D Scenes

2025-10-25 · Kunal Bhosikar, Siddharth Katageri, Vivek Madhavaram, Kai Han 외 arxiv

Generating realistic full-body motion interacting with objects is critical for applications in robotics, virtual reality, and human-computer interaction. While existing methods can generate full-body motion within 3D sce…

WonderWorld: Interactive 3D Scene Generation from a Single Image

2024-06-13 · CVPR 2025 1 · Hong-Xing Yu, Haoyi Duan, Charles Herrmann, William T. Freeman 외

We present WonderWorld, a novel framework for interactive 3D scene generation that enables users to interactively specify scene contents and layout and see the created scenes in low latency. The major challenge lies in a…

Depth EstimationGPUNavigateScene Generation