paper-with-me

홈 › Papers

MVG4D: Image Matrix-Based Multi-View and Motion Generation for 4D Content Creation from a Single Image

2025-07-24 · DongFu Yin, Xiaotian Chen, Fei Richard Yu, Xuanchen Li, Xinhao Zhang arxiv

Advances in generative modeling have significantly enhanced digital content creation, extending from 2D images to complex 3D and 4D scenes. Despite substantial progress, producing high-fidelity and temporally consistent dynamic 4D content remains a challenge. In this paper, we propose MVG4D, a novel framework that generates dynamic 4D content from a single still image by combining multi-view synthesis with 4D Gaussian Splatting (4D GS). At its core, MVG4D employs an image matrix module that synthesizes temporally coherent and spatially diverse multi-view images, providing rich supervisory signals for downstream 3D and 4D reconstruction. These multi-view images are used to optimize a 3D Gaussian point cloud, which is further extended into the temporal domain via a lightweight deformation network. Our method effectively enhances temporal consistency, geometric fidelity, and visual realism, addressing key challenges in motion discontinuity and background degradation that affect prior 4D GS-based methods. Extensive experiments on the Objaverse dataset demonstrate that MVG4D outperforms state-of-the-art baselines in CLIP-I, PSNR, FVD, and time efficiency. Notably, it reduces flickering artifacts and sharpens structural details across views and time, enabling more immersive AR/VR experiences. MVG4D sets a new direction for efficient and controllable 4D generation from minimal inputs.

📄 PDF Abstract BibTeX arXiv:2507.18371

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Multi-view Registration Based on Weighted Low Rank and Sparse Matrix Decomposition of Motions

2017-09-25 · Congcong Jin, Jihua Zhu, Yaochen Li, Shanmin Pang 외

Recently, the low rank and sparse (LRS) matrix decomposition has been introduced as an effective mean to solve the multi-view registration. It views each available relative motion as a block element to reconstruct one ma…

A New Rank Constraint on Multi-view Fundamental Matrices, and its Application to Camera Location Recovery

2017-02-10 · CVPR 2017 7 · Soumyadip Sengupta, Tal Amir, Meirav Galun, Tom Goldstein 외

Accurate estimation of camera matrices is an important step in structure from motion algorithms. In this paper we introduce a novel rank constraint on collections of fundamental matrices in multi-view settings. We show t…

Reangle-A-Video: 4D Video Generation as Video-to-Video Translation

2025-03-12 · Hyeonho Jeong, Suhyeon Lee, Jong Chul Ye

We introduce Reangle-A-Video, a unified framework for generating synchronized multi-view videos from a single input video. Unlike mainstream approaches that train multi-view video diffusion models on large-scale 4D datas…

TranslationVideo Generation

Articulate That Object Part (ATOP): 3D Part Articulation from Text and Motion Personalization

2025-02-11 · Aditya Vora, Sauradip Nag, Hao Zhang

We present ATOP (Articulate That Object Part), a novel method based on motion personalization to articulate a 3D object with respect to a part and its motion as prescribed in a text prompt. Specifically, the text input a…

Image GenerationMotion GenerationObjectVideo Generation

PV3D: A 3D Generative Model for Portrait Video Generation

2022-12-13 · Zhongcong Xu, Jianfeng Zhang, Jun Hao Liew, Wenqing Zhang 외

Recent advances in generative adversarial networks (GANs) have demonstrated the capabilities of generating stunning photo-realistic portrait images. While some prior works have applied such image GANs to unconditional 2D…

Video Generation