Semantics-Driven Unsupervised Learning for Monocular Depth and Ego-Motion Estimation
We propose a semantics-driven unsupervised learning approach for monocular depth and ego-motion estimation from videos in this paper. Recent unsupervised learning methods employ photometric errors between synthetic view and actual image as a supervision signal for training. In our method, we exploit semantic segmentation information to mitigate the effects of dynamic objects and occlusions in the scene, and to improve depth prediction performance by considering the correlation between depth and semantics. To avoid costly labeling process, we use noisy semantic segmentation results obtained by a pre-trained semantic segmentation network. In addition, we minimize the position error between the corresponding points of adjacent frames to utilize 3D spatial information. Experimental results on the KITTI dataset show that our method achieves good performance in both depth and ego-motion estimation tasks.
Code (0)
등록된 구현이 없습니다.
Tasks
Depth EstimationDepth PredictionMotion EstimationPositionSegmentationSemantic SegmentationSimilar Papers 제목 키워드 기반
Unsupervised Monocular Depth and Ego-motion Learning with Structure and Semantics
We present an approach which takes advantage of both structure and semantics for unsupervised monocular learning of depth and ego-motion. More specifically, we model the motion of individual objects and learn their 3D mo…
Depth And Camera MotionDepth EstimationMonocular Depth EstimationMotion EstimationDynamo-Depth: Fixing Unsupervised Depth Estimation for Dynamical Scenes
Unsupervised monocular depth estimation techniques have demonstrated encouraging results but typically assume that the scene is static. These techniques suffer when trained on dynamical scenes, where apparent object moti…
Depth EstimationMonocular Depth EstimationMotion SegmentationSegmentation+1Geometry meets semantics for semi-supervised monocular depth estimation
Depth estimation from a single image represents a very exciting challenge in computer vision. While other image-based depth sensing techniques leverage on the geometry between different viewpoints (e.g., stereo or struct…
DecoderDepth EstimationDepth PredictionMonocular Depth Estimation+1DeFeat-Net: General Monocular Depth via Simultaneous Unsupervised Representation Learning
In the current monocular depth research, the dominant approach is to employ unsupervised training on large datasets, driven by warped photometric consistency. Such approaches lack robustness and are unable to generalize …
Depth EstimationMonocular Depth EstimationRepresentation LearningRM-Depth: Unsupervised Learning of Recurrent Monocular Depth in Dynamic Scenes
Unsupervised methods have showed promising results on monocular depth estimation. However, the training data must be captured in scenes without moving objects. To push the envelope of accuracy, recent methods tend to inc…
DecoderDepth EstimationMonocular Depth EstimationUnsupervised Monocular Depth Estimation