paper-with-me

홈 › Papers

MegaDepth: Learning Single-View Depth Prediction from Internet Photos

2018-04-02 · CVPR 2018 6 · Zhengqi Li, Noah Snavely

Single-view depth prediction is a fundamental problem in computer vision. Recently, deep learning methods have led to significant progress, but such methods are limited by the available training data. Current datasets based on 3D sensors have key limitations, including indoor-only images (NYU), small numbers of training examples (Make3D), and sparse sampling (KITTI). We propose to use multi-view Internet photo collections, a virtually unlimited data source, to generate training data via modern structure-from-motion and multi-view stereo (MVS) methods, and present a large depth dataset called MegaDepth based on this idea. Data derived from MVS comes with its own challenges, including noise and unreconstructable objects. We address these challenges with new data cleaning methods, as well as automatically augmenting our data with ordinal depth relations generated using semantic segmentation. We validate the use of large amounts of Internet data by showing that models trained on MegaDepth exhibit strong generalization-not only to novel scenes, but also to other diverse datasets including Make3D, KITTI, and DIW, even when no images from those datasets are seen during training.

📄 PDF Abstract BibTeX arXiv:1804.00607

Code (2)

fabio-sim/DeDoDe-ONNX-TensorRT pytorch
zhengqili/MegaDepth pytorch

Tasks

Depth EstimationDepth PredictionPredictionSemantic Segmentation

Similar Papers 제목 키워드 기반

Long-tail Internet photo reconstruction

2026-04-24 · Yuan Li, Yuanbo Xiangli, Hadar Averbuch-Elor, Noah Snavely 외 arxiv

Internet photo collections exhibit an extremely long-tailed distribution: a few famous landmarks are densely photographed and easily reconstructed in 3D, while most real-world sites are represented with sparse, noisy, un…

AerialMegaDepth: Learning Aerial-Ground Reconstruction and View Synthesis

2025-04-17 · CVPR 2025 1 · Khiem Vuong, Anurag Ghosh, Deva Ramanan, Srinivasa Narasimhan 외

We explore the task of geometric reconstruction of images captured from a mixture of ground and aerial views. Current state-of-the-art learning-based approaches fail to handle the extreme viewpoint variation between aeri…

Novel View Synthesis

360 in the Wild: Dataset for Depth Prediction and View Synthesis

2024-06-27 · Kibaek Park, Francois Rameau, Jaesik Park, In So Kweon

The large abundance of perspective camera datasets facilitated the emergence of novel learning-based strategies for various tasks, such as camera localization, single image depth estimation, or view synthesis. However, p…

Camera LocalizationDepth EstimationDepth Prediction

Symmetry-aware Depth Estimation using Deep Neural Networks

2016-04-20 · Guilin Liu, Chao Yang, Zimo Li, Duygu Ceylan 외

Due to the abundance of 2D product images from the Internet, developing efficient and scalable algorithms to recover the missing depth information is central to many applications. Recent works have addressed the single-v…

Depth Estimation

Learning Single-Image Depth from Videos using Quality Assessment Networks

2018-06-25 · CVPR 2019 6 · Weifeng Chen, Shengyi Qian, Jia Deng

Depth estimation from a single image in the wild remains a challenging problem. One main obstacle is the lack of high-quality training data for images in the wild. In this paper we propose a method to automatically gener…

Depth Estimation