paper-with-me

홈 › Papers

Generating 3D-Consistent Videos from Unposed Internet Photos

2024-11-20 · CVPR 2025 1 · Gene Chou, Kai Zhang, Sai Bi, Hao Tan, Zexiang Xu, Fujun Luan, Bharath Hariharan, Noah Snavely

We address the problem of generating videos from unposed internet photos. A handful of input images serve as keyframes, and our model interpolates between them to simulate a path moving between the cameras. Given random images, a model's ability to capture underlying geometry, recognize scene identity, and relate frames in terms of camera position and orientation reflects a fundamental understanding of 3D structure and scene layout. However, existing video models such as Luma Dream Machine fail at this task. We design a self-supervised method that takes advantage of the consistency of videos and variability of multiview internet photos to train a scalable, 3D-aware video model without any 3D annotations such as camera parameters. We validate that our method outperforms all baselines in terms of geometric and appearance consistency. We also show our model benefits applications that enable camera control, such as 3D Gaussian Splatting. Our results suggest that we can scale up scene-level 3D learning using only 2D data such as videos and multiview internet photos.

📄 PDF Abstract BibTeX arXiv:2411.13549

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Neural Rerendering in the Wild

2019-04-08 · CVPR 2019 6 · Moustafa Meshry, Dan B. Goldman, Sameh Khamis, Hugues Hoppe 외

We explore total scene capture -- recording, modeling, and rerendering a scene under varying appearance such as season and time of day. Starting from internet photos of a tourist landmark, we apply traditional 3D reconst…

3D Reconstruction

Pose-Free Omnidirectional Gaussian Splatting for 360-Degree Videos with Consistent Depth Priors

2026-03-24 · Chuanqing Zhuang, Xin Lu, Zehui Deng, Zhengda Lu 외 arxiv

Omnidirectional 3D Gaussian Splatting with panoramas is a key technique for 3D scene representation, and existing methods typically rely on slow SfM to provide camera poses and sparse points priors. In this work, we prop…

Camera Pose EstimationNovel View Synthesis

StegaStamp: Invisible Hyperlinks in Physical Photographs

2019-04-10 · CVPR 2020 6 · Matthew Tancik, Ben Mildenhall, Ren Ng

Printed and digitally displayed photos have the ability to hide imperceptible digital data that can be accessed through internet-connected imaging systems. Another way to think about this is physical photographs that hav…

Steganographics

LucidFusion: Generating 3D Gaussians with Arbitrary Unposed Images

2024-10-21 · Hao He, Yixun Liang, Luozhou Wang, Yuanhao Cai 외

Recent large reconstruction models have made notable progress in generating high-quality 3D objects from single images. However, these methods often struggle with controllability, as they lack information from multiple v…

3D GenerationImage to 3D

Head Reconstruction from Internet Photos

2018-09-13 · Shu Liang, Linda G. Shapiro, Ira Kemelmacher-Shlizerman

3D face reconstruction from Internet photos has recently produced exciting results. A person's face, e.g., Tom Hanks, can be modeled and animated in 3D from a completely uncalibrated photo collection. Most methods, howev…

3D Face ReconstructionFace Reconstruction