paper-with-me

홈 › Papers

Epipolar Geometry based Learning of Multi-view Depth and Ego-Motion from Monocular Sequences

2018-12-23 · Vignesh Prasad, Dipanjan Das, Brojeshwar Bhowmick

Deep approaches to predict monocular depth and ego-motion have grown in recent years due to their ability to produce dense depth from monocular images. The main idea behind them is to optimize the photometric consistency over image sequences by warping one view into another, similar to direct visual odometry methods. One major drawback is that these methods infer depth from a single view, which might not effectively capture the relation between pixels. Moreover, simply minimizing the photometric loss does not ensure proper pixel correspondences, which is a key factor for accurate depth and pose estimations. In contrast, we propose a 2-view depth network to infer the scene depth from consecutive frames, thereby learning inter-pixel relationships. To ensure better correspondences, thereby better geometric understanding, we propose incorporating epipolar constraints to make the learning more geometrically sound. We use the Essential matrix obtained using Nist'er's Five Point Algorithm, to enforce meaningful geometric constraints, rather than using it as training labels. This allows us to use lesser no. of trainable parameters compared to state-of-the-art methods. The proposed method results in better depth images and pose estimates, which capture the scene structure and motion in a better way. Such a geometrically constrained learning performs successfully even in cases where simply minimizing the photometric error would fail.

📄 PDF Abstract BibTeX arXiv:1812.11922

Code (0)

등록된 구현이 없습니다.

Tasks

Visual Odometry

Similar Papers 제목 키워드 기반

DGSfM: Depth-Guided Scale-Aware Global Structure-from-Motion

2026-07-10 · Sithu Aung, Viktor Kocur, Yaqing Ding, Torsten Sattler 외 arxiv

Global Structure-from-Motion (SfM) is an efficient paradigm for recovering camera poses and sparse 3D structure from unordered images. However, its reliance on scale-ambiguous epipolar geometry makes global positioning s…

DualRefine: Self-Supervised Depth and Pose Estimation Through Iterative Epipolar Sampling and Refinement Toward Equilibrium

2023-04-07 · CVPR 2023 1 · Antyanta Bangunharcana, Ahmed Magd, Kyung-Soo Kim

Self-supervised multi-frame depth estimation achieves high accuracy by computing matching costs of pixel correspondences between adjacent frames, injecting geometric information into the network. These pixel-corresponden…

Depth EstimationDepth PredictionPose Estimation

DELS-MVS: Deep Epipolar Line Search for Multi-View Stereo

2022-12-13 · Christian Sormann, Emanuele Santellani, Mattia Rossi, Andreas Kuhn 외

We propose a novel approach for deep learning-based Multi-View Stereo (MVS). For each pixel in the reference image, our method leverages a deep architecture to search for the corresponding point in the source image direc…

Camera Calibration from Dynamic Silhouettes Using Motion Barcodes

2015-06-25 · CVPR 2016 6 · Gil Ben-Artzi, Yoni Kasten, Shmuel Peleg, Michael Werman

Computing the epipolar geometry between cameras with very different viewpoints is often problematic as matching points are hard to find. In these cases, it has been proposed to use information from dynamic objects in the…

Camera Calibration

Novel View Synthesis of Dynamic Scenes with Globally Coherent Depths from a Monocular Camera

2020-04-02 · CVPR 2020 6 · Jae Shin Yoon, Kihwan Kim, Orazio Gallo, Hyun Soo Park 외

This paper presents a new method to synthesize an image from arbitrary views and times given a collection of images of a dynamic scene. A key challenge for the novel view synthesis arises from dynamic scene reconstructio…

Depth EstimationNovel View Synthesis