paper-with-me

홈 › Papers

Synthetic-to-Real Self-supervised Robust Depth Estimation via Learning with Motion and Structure Priors

2025-03-26 · CVPR 2025 1 · Weilong Yan, Ming Li, Haipeng Li, Shuwei Shao, Robby T. Tan

Self-supervised depth estimation from monocular cameras in diverse outdoor conditions, such as daytime, rain, and nighttime, is challenging due to the difficulty of learning universal representations and the severe lack of labeled real-world adverse data. Previous methods either rely on synthetic inputs and pseudo-depth labels or directly apply daytime strategies to adverse conditions, resulting in suboptimal results. In this paper, we present the first synthetic-to-real robust depth estimation framework, incorporating motion and structure priors to capture real-world knowledge effectively. In the synthetic adaptation, we transfer motion-structure knowledge inside cost volumes for better robust representation, using a frozen daytime model to train a depth estimator in synthetic adverse conditions. In the innovative real adaptation, which targets to fix synthetic-real gaps, models trained earlier identify the weather-insensitive regions with a designed consistency-reweighting strategy to emphasize valid pseudo-labels. We introduce a new regularization by gathering explicit depth distributions to constrain the model when facing real-world data. Experiments show that our method outperforms the state-of-the-art across diverse conditions in multi-frame and single-frame evaluations. We achieve improvements of 7.5% and 4.3% in AbsRel and RMSE on average for nuScenes and Robotcar datasets (daytime, nighttime, rain). In zero-shot evaluation of DrivingStereo (rain, fog), our method generalizes better than the previous ones.

📄 PDF Abstract BibTeX arXiv:2503.20211

Code (1)

davidyan2001/synthetic2real-depth 공식 구현 pytorch

Tasks

Depth EstimationWorld Knowledge

Similar Papers 제목 키워드 기반

$S^3$Net: Semantic-Aware Self-supervised Depth Estimation with Monocular Videos and Synthetic Data

2020-07-28 · Bin Cheng, Inderjot Singh Saggu, Raunak Shah, Gaurav Bansal 외

Solving depth estimation with monocular cameras enables the possibility of widespread use of cameras as low-cost depth estimation sensors in applications such as autonomous driving and robotics. However, learning such a …

Autonomous DrivingDepth EstimationDomain AdaptationPanoptic Segmentation

S³Net: Semantic-Aware Self-supervised Depth Estimation with Monocular Videos and Synthetic Data

2020-08-01 · ECCV 2020 8 · Bin Cheng, Inderjot Singh Saggu, Raunak Shah, Gaurav Bansal 외

Solving depth estimation with monocular cameras enables the possibility of widespread use of cameras as low-cost depth estimation sensors in applications such as autonomous driving and robotics. In order to learn such a …

Autonomous DrivingDepth EstimationDomain Adaptation

Self-Supervised Learning of Domain Invariant Features for Depth Estimation

2021-06-04 · Hiroyasu Akada, Shariq Farooq Bhat, Ibraheem Alhashim, Peter Wonka

We tackle the problem of unsupervised synthetic-to-real domain adaptation for single image depth estimation. An essential building block of single image depth estimation is an encoder-decoder task network that takes RGB …

Depth EstimationDomain AdaptationImage-to-Image TranslationRepresentation Learning+3

Structure-preserving Image Translation for Depth Estimation in Colonoscopy Video

2024-08-19 · Shuxian Wang, Akshay Paruchuri, Zhaoxi Zhang, Sarah McGill 외

Monocular depth estimation in colonoscopy video aims to overcome the unusual lighting properties of the colonoscopic environment. One of the major challenges in this area is the domain gap between annotated but unrealist…

Depth EstimationMonocular Depth EstimationTranslation

Self-Supervised Spatially Variant PSF Estimation for Aberration-Aware Depth-from-Defocus

2024-02-28 · Zhuofeng Wu, Yusuke Monno, Masatoshi Okutomi

In this paper, we address the task of aberration-aware depth-from-defocus (DfD), which takes account of spatially variant point spread functions (PSFs) of a real camera. To effectively obtain the spatially variant PSFs o…

Depth EstimationSelf-Supervised Learning