paper-with-me

Papers

Versatile Depth Estimator Based on Common Relative Depth Estimation and Camera-Specific Relative-to-Metric Depth Conversion

2023-03-20 · Jinyoung Jun, Jae-Han Lee, Chang-Su Kim

A typical monocular depth estimator is trained for a single camera, so its performance drops severely on images taken with different cameras. To address this issue, we propose a versatile depth estimator (VDE), composed of a common relative depth estimator (CRDE) and multiple relative-to-metric converters (R2MCs). The CRDE extracts relative depth information, and each R2MC converts the relative information to predict metric depths for a specific camera. The proposed VDE can cope with diverse scenes, including both indoor and outdoor scenes, with only a 1.12\% parameter increase per camera. Experimental results demonstrate that VDE supports multiple cameras effectively and efficiently and also achieves state-of-the-art performance in the conventional single-camera scenario.

📄 PDF Abstract BibTeX arXiv:2303.10991

Code (0)

등록된 구현이 없습니다.

Tasks

Depth Estimation

Similar Papers 제목 키워드 기반

RSA: Resolving Scale Ambiguities in Monocular Depth Estimators through Language Descriptions

2024-10-03 · Ziyao Zeng, Yangchao Wu, Hyoungseob Park, Daniel Wang 외

We propose a method for metric-scale monocular depth estimation. Inferring depth from a single image is an ill-posed problem due to the loss of scale from perspective projection during the image formation process. Any sc…

Depth EstimationMonocular Depth Estimation

SPADE: Sparsity Adaptive Depth Estimator for Zero-Shot, Real-Time, Monocular Depth Estimation in Underwater Environments

2025-10-29 · Hongjie Zhang, Gideon Billings, Stefan B. Williams arxiv

Underwater infrastructure requires frequent inspection and maintenance due to harsh marine conditions. Current reliance on human divers or remotely operated vehicles is limited by perceptual and operational challenges, e…

Monocular Depth Estimation

Compose and Conquer: Diffusion-Based 3D Depth Aware Composable Image Synthesis

2024-01-17 · Jonghyun Lee, Hansam Cho, Youngjoon Yoo, Seoung Bum Kim 외

Addressing the limitations of text as a source of accurate layout representation in text-conditional diffusion models, many works incorporate additional signals to condition certain attributes within a generated image. A…

DisentanglementImage Generation

FlashDepth: Real-time Streaming Video Depth Estimation at 2K Resolution

2025-04-09 · Gene Chou, Wenqi Xian, Guandao Yang, Mohamed Abdelfattah 외

A versatile video depth estimation model should (1) be accurate and consistent across frames, (2) produce high-resolution depth maps, and (3) support real-time streaming. We propose FlashDepth, a method that satisfies al…

2kDecision MakingDepth EstimationVideo Editing

Repurposing Diffusion-Based Image Generators for Monocular Depth Estimation

2023-12-04 · CVPR 2024 1 · Bingxin Ke, Anton Obukhov, Shengyu Huang, Nando Metzger 외

Monocular depth estimation is a fundamental computer vision task. Recovering 3D depth from a single image is geometrically ill-posed and requires scene understanding, so it is not surprising that the rise of deep learnin…

Depth EstimationGPUMonocular Depth EstimationScene Understanding+1