paper-with-me

홈 › Papers

Gaussian Variation Field Diffusion for High-fidelity Video-to-4D Synthesis

2025-07-31 · Bowen Zhang, Sicheng Xu, Chuxin Wang, Jiaolong Yang, Feng Zhao, Dong Chen, Baining Guo arxiv

In this paper, we present a novel framework for video-to-4D generation that creates high-quality dynamic 3D content from single video inputs. Direct 4D diffusion modeling is extremely challenging due to costly data construction and the high-dimensional nature of jointly representing 3D shape, appearance, and motion. We address these challenges by introducing a Direct 4DMesh-to-GS Variation Field VAE that directly encodes canonical Gaussian Splats (GS) and their temporal variations from 3D animation data without per-instance fitting, and compresses high-dimensional animations into a compact latent space. Building upon this efficient representation, we train a Gaussian Variation Field diffusion model with temporal-aware Diffusion Transformer conditioned on input videos and canonical GS. Trained on carefully-curated animatable 3D objects from the Objaverse dataset, our model demonstrates superior generation quality compared to existing methods. It also exhibits remarkable generalization to in-the-wild video inputs despite being trained exclusively on synthetic data, paving the way for generating high-quality animated 3D content. Project page: https://gvfdiffusion.github.io/.

📄 PDF Abstract BibTeX arXiv:2507.23785

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

ProDiG: Progressive Diffusion-Guided Gaussian Splatting for Aerial to Ground Reconstruction

2026-04-02 · Sirshapan Mitra, Yogesh S. Rawat arxiv

Generating ground-level views and coherent 3D site models from aerial-only imagery is challenging due to extreme viewpoint changes, missing intermediate observations, and large scale variations. Existing methods either r…

Atlas Gaussians Diffusion for 3D Generation

2024-08-23 · Haitao Yang, Yuan Dong, Hanwen Jiang, Dejia Xu 외

Using the latent diffusion model has proven effective in developing novel 3D generation techniques. To harness the latent diffusion model, a key challenge is designing a high-fidelity and efficient representation that li…

3D Generation

FreeFix: Boosting 3D Gaussian Splatting via Fine-Tuning-Free Diffusion Models

2026-01-28 · Hongyu Zhou, Zisen Shao, Sheng Miao, Pan Wang 외 arxiv

Neural Radiance Fields and 3D Gaussian Splatting have advanced novel view synthesis, yet still rely on dense inputs and often degrade at extrapolated views. Recent approaches leverage generative models, such as diffusion…

Novel View Synthesis

GPAvatar: High-fidelity Head Avatars by Learning Efficient Gaussian Projections

2025-01-01 · CVPR 2025 1 · Wei-Qi Feng, Dong Han, Ze-Kang Zhou, Shunkai Li 외

Existing radiance field-based head avatar methods have mostly relied on pre-computed explicit priors (e.g., mesh, point) or neural implicit representations, making it challenging to achieve high fidelity with both co…

Computational Efficiency

PiRD: Physics-informed Residual Diffusion for Flow Field Reconstruction

2024-04-12 · Siming Shan, Pengkai Wang, Song Chen, Jiaxu Liu 외

The use of machine learning in fluid dynamics is becoming more common to expedite the computation when solving forward and inverse problems of partial differential equations. Yet, a notable challenge with existing convol…