Guiding Monocular Depth Estimation Using Depth-Attention Volume
Recovering the scene depth from a single image is an ill-posed problem that requires additional priors, often referred to as monocular depth cues, to disambiguate different 3D interpretations. In recent works, those priors have been learned in an end-to-end manner from large datasets by using deep neural networks. In this paper, we propose guiding depth estimation to favor planar structures that are ubiquitous especially in indoor environments. This is achieved by incorporating a non-local coplanarity constraint to the network with a novel attention mechanism called depth-attention volume (DAV). Experiments on two popular indoor datasets, namely NYU-Depth-v2 and ScanNet, show that our method achieves state-of-the-art depth estimation results while using only a fraction of the number of parameters needed by the competing methods.
Code (2)
Tasks
Depth EstimationMonocular Depth EstimationSimilar Papers 제목 키워드 기반
Depth-Relative Self Attention for Monocular Depth Estimation
Monocular depth estimation is very challenging because clues to the exact depth are incomplete in a single RGB image. To overcome the limitation, deep neural networks rely on various visual hints such as size, shade, and…
Depth EstimationMonocular Depth EstimationLook Deeper into Depth: Monocular Depth Estimation with Semantic Booster and Attention-Driven Loss
Monocular depth estimation benefits greatly from learning based techniques. By studying the training data, we observe that the per-pixel depth values in existing datasets typically exhibit a long-tailed distribution. How…
Depth EstimationMonocular Depth EstimationMAMo: Leveraging Memory and Attention for Monocular Video Depth Estimation
We propose MAMo, a novel memory and attention frame-work for monocular video depth estimation. MAMo can augment and improve any single-image depth estimation networks into video depth estimation models, enabling them to …
Depth EstimationDepth PredictionMonocular Depth EstimationManydepth2: Motion-Aware Self-Supervised Multi-Frame Monocular Depth Estimation in Dynamic Scenes
Despite advancements in self-supervised monocular depth estimation, challenges persist in dynamic scenarios due to the dependence on assumptions about a static world. In this paper, we present Manydepth2, to achieve prec…
Camera Pose EstimationComputational EfficiencyDepth EstimationMonocular Depth Estimation+1Structure-Attentioned Memory Network for Monocular Depth Estimation
Monocular depth estimation is a challenging task that aims to predict a corresponding depth map from a given single RGB image. Recent deep learning models have been proposed to predict the depth from the image by learnin…
Depth EstimationDomain AdaptationMonocular Depth Estimation