paper-with-me

Papers

Unsupervised Scale-consistent Depth Learning from Video

2021-05-25 · Jia-Wang Bian, Huangying Zhan, Naiyan Wang, Zhichao Li, Le Zhang, Chunhua Shen, Ming-Ming Cheng, Ian Reid

We propose a monocular depth estimator SC-Depth, which requires only unlabelled videos for training and enables the scale-consistent prediction at inference time. Our contributions include: (i) we propose a geometry consistency loss, which penalizes the inconsistency of predicted depths between adjacent views; (ii) we propose a self-discovered mask to automatically localize moving objects that violate the underlying static scene assumption and cause noisy signals during training; (iii) we demonstrate the efficacy of each component with a detailed ablation study and show high-quality depth estimation results in both KITTI and NYUv2 datasets. Moreover, thanks to the capability of scale-consistent prediction, we show that our monocular-trained deep networks are readily integrated into the ORB-SLAM2 system for more robust and accurate tracking. The proposed hybrid Pseudo-RGBD SLAM shows compelling results in KITTI, and it generalizes well to the KAIST dataset without additional training. Finally, we provide several demos for qualitative evaluation.

📄 PDF Abstract BibTeX arXiv:2105.11610

Code (2)

JiawangBian/sc_depth_pl 공식 구현 pytorch
JiawangBian/SC-SfMLearner-Release pytorch

Tasks

Depth EstimationMonocular Depth EstimationMonocular Visual OdometrySimultaneous Localization and Mapping

Similar Papers 제목 키워드 기반

Unsupervised Scale-consistent Depth and Ego-motion Learning from Monocular Video

2019-08-28 · NeurIPS 2019 12 · Jia-Wang Bian, Zhichao Li, Naiyan Wang, Huangying Zhan 외

Recent work has shown that CNN-based depth and ego-motion estimators can be learned using unlabelled monocular videos. However, the performance is limited by unidentified moving objects that violate the underlying static…

Camera Pose EstimationDepth And Camera MotionDepth EstimationMonocular Depth Estimation+1

DesNet: Decomposed Scale-Consistent Network for Unsupervised Depth Completion

2022-11-20 · Zhiqiang Yan, Kun Wang, Xiang Li, Zhenyu Zhang 외

Unsupervised depth completion aims to recover dense depth from the sparse one without using the ground-truth annotation. Although depth measurement obtained from LiDAR is usually sparse, it contains valid and real distan…

Depth CompletionDepth EstimationDepth Predictionvalid

DepthSync: Diffusion Guidance-Based Depth Synchronization for Scale- and Geometry-Consistent Video Depth Estimation

2025-07-02 · Yue-Jiang Dong, Wang Zhao, Jiale Xu, Ying Shan 외 arxiv

Diffusion-based video depth estimation methods have achieved remarkable success with strong generalization ability. However, predicting depth for long videos remains challenging. Existing methods typically split videos i…

Depth Estimation

Scene-Centric Unsupervised Video Panoptic Segmentation

2026-06-03 · Christoph Reich, Oliver Hahn, Nikita Araslanov, Laura Leal-Taixé 외 arxiv

Video panoptic segmentation (VPS) aims to jointly detect, segment, and track all objects while partitioning the video into semantically consistent regions. We introduce the task setting of unsupervised VPS, omitting any …

Video Panoptic SegmentationScene UnderstandingVideo SegmentationImage Segmentation

Unsupervised Learning of Depth and Ego-Motion from Monocular Video Using 3D Geometric Constraints

2018-02-15 · CVPR 2018 6 · Reza Mahjourian, Martin Wicke, Anelia Angelova

We present a novel approach for unsupervised learning of depth and ego-motion from monocular video. Unsupervised learning removes the need for separate supervisory signals (depth or ego-motion ground truth, or multi-view…

3D geometryDepth And Camera MotionDepth EstimationMonocular Depth Estimation+1