Deep Robust Single Image Depth Estimation Neural Network Using Scene Understanding
Single image depth estimation (SIDE) plays a crucial role in 3D computer vision. In this paper, we propose a two-stage robust SIDE framework that can perform blind SIDE for both indoor and outdoor scenes. At the first stage, the scene understanding module will categorize the RGB image into different depth-ranges. We introduce two different scene understanding modules based on scene classification and coarse depth estimation respectively. At the second stage, SIDE networks trained by the images of specific depth-range are applied to obtain an accurate depth map. In order to improve the accuracy, we further design a multi-task encoding-decoding SIDE network DS-SIDENet based on depthwise separable convolutions. DS-SIDENet is optimized to minimize both depth classification and depth regression losses. This improves the accuracy compared to a single-task SIDE network. Experimental results demonstrate that training DS-SIDENet on an individual dataset such as NYU achieves competitive performance to the state-of-art methods with much better efficiency. Ours proposed robust SIDE framework also shows good performance for the ScanNet indoor images and KITTI outdoor images simultaneously. It achieves the top performance compared to the Robust Vision Challenge (ROB) 2018 submissions.
Code (0)
등록된 구현이 없습니다.
Tasks
Depth EstimationGeneral ClassificationScene ClassificationScene UnderstandingSimilar Papers 제목 키워드 기반
To complete or to estimate, that is the question: A Multi-Task Approach to Depth Completion and Monocular Depth Estimation
Robust three-dimensional scene understanding is now an ever-growing area of research highly relevant in many real-world applications such as autonomous driving and robotic navigation. In this paper, we propose a multi-ta…
Autonomous DrivingDepth CompletionDepth EstimationMonocular Depth Estimation+2Single Image Depth Estimation: An Overview
We review solutions to the problem of depth estimation, arguably the most important subtask in scene understanding. We focus on the single image depth estimation problem. Due to its properties, the single image depth est…
Deep LearningDepth EstimationScene UnderstandingSemantic Segmentation+1MGNet: Monocular Geometric Scene Understanding for Autonomous Driving
We introduce MGNet, a multi-task framework for monocular geometric scene understanding. We define monocular geometric scene understanding as the combination of two known tasks: Panoptic segmentation and self-supervised m…
Autonomous DrivingDepth EstimationGPUMonocular Depth Estimation+2Multi-task Geometric Estimation of Depth and Surface Normal from Monocular 360° Images
Geometric estimation is required for scene understanding and analysis in panoramic 360{\deg} images. Current methods usually predict a single feature, such as depth or surface normal. These methods can lack robustness, e…
Multi-Task LearningScene UnderstandingSurface Normal EstimationIndoor Scene Structure Analysis for Single Image Depth Estimation
We tackle the problem of single image depth estimation, which, without additional knowledge, suffers from many ambiguities. Unlike previous approaches that only reason locally, we propose to exploit the global structure …
Depth Estimation