paper-with-me

홈 › Papers

GaussFusion: Improving 3D Reconstruction in the Wild with A Geometry-Informed Video Generator

2026-03-26 · Liyuan Zhu, Manjunath Narayana, Michal Stary, Will Hutchcroft, Gordon Wetzstein, Iro Armeni arxiv

We present GaussFusion, a novel approach for improving 3D Gaussian splatting (3DGS) reconstructions in the wild through geometry-informed video generation. GaussFusion mitigates common 3DGS artifacts, including floaters, flickering, and blur caused by camera pose errors, incomplete coverage, and noisy geometry initialization. Unlike prior RGB-based approaches limited to a single reconstruction pipeline, our method introduces a geometry-informed video-to-video generator that refines 3DGS renderings across both optimization-based and feed-forward methods. Given an existing reconstruction, we render a Gaussian primitive video buffer encoding depth, normals, opacity, and covariance, which the generator refines to produce temporally coherent, artifact-free frames. We further introduce an artifact synthesis pipeline that simulates diverse degradation patterns, ensuring robustness and generalization. GaussFusion achieves state-of-the-art performance on novel-view synthesis benchmarks, and an efficient variant runs in real time at 15 FPS while maintaining similar performance, enabling interactive 3D applications.

📄 PDF Abstract BibTeX arXiv:2603.25053

Code (0)

등록된 구현이 없습니다.

Tasks

3D ReconstructionVideo Generation

Similar Papers 제목 키워드 기반

GaussFusion: Towards Multimodal 3D Gaussian Pretraining

2026-07-07 · Zhixuan You, Jihua Zhu, Yiding Sun, Zihao Guo 외 arxiv

3D Gaussian Splatting provides an explicit representation that jointly models geometry and appearance, serving as a scalable foundation for 3D representation learning. Existing pre-training methods for Gaussian represent…

Representation Learning

Lift4D: Harmonizing Single-View 3D Estimation for 4D Reconstruction In-the-Wild

2026-06-22 · Yehonathan Litman, Xiaoxuan Ma, Manan Shah, Nicolas Ugrinovic 외 arxiv

Reconstructing dynamic non-rigid objects from monocular video requires integrating visual cues from direct observations with data-driven priors over geometry and appearance. Prior approaches either learn to directly pred…

Single-View 3D Reconstruction

CRISP: Contact-Guided Real2Sim from Monocular Video with Planar Scene Primitives

2025-12-16 · Zihan Wang, Jiashun Wang, Jeff Tan, Yiwen Zhao 외 arxiv

We introduce CRISP, a method that recovers simulatable human motion and scene geometry from monocular video. Prior work on joint human-scene reconstruction relies on data-driven priors and joint optimization with no phys…

Reinforcement Learning

Vid2Avatar: 3D Avatar Reconstruction from Videos in the Wild via Self-supervised Scene Decomposition

2023-02-22 · CVPR 2023 1 · Chen Guo, Tianjian Jiang, Xu Chen, Jie Song 외

We present Vid2Avatar, a method to learn human avatars from monocular in-the-wild videos. Reconstructing humans that move naturally from monocular in-the-wild videos is difficult. Solving it requires accurately separatin…

3D Human Reconstructionglobal-optimizationSurface Reconstruction

Joint Optimization for 4D Human-Scene Reconstruction in the Wild

2025-01-04 · Zhizheng Liu, Joe Lin, Wayne Wu, Bolei Zhou

Reconstructing human motion and its surrounding environment is crucial for understanding human-scene interaction and predicting human movements in the scene. While much progress has been made in capturing human-scene int…

Human Mesh RecoveryMotion Estimation