paper-with-me

Papers

GVDepth: Zero-Shot Monocular Depth Estimation for Ground Vehicles based on Probabilistic Cue Fusion

2024-12-08 · Karlo Koledic, Luka Petrovic, Ivan Markovic, Ivan Petrovic

Generalizing metric monocular depth estimation presents a significant challenge due to its ill-posed nature, while the entanglement between camera parameters and depth amplifies issues further, hindering multi-dataset training and zero-shot accuracy. This challenge is particularly evident in autonomous vehicles and mobile robotics, where data is collected with fixed camera setups, limiting the geometric diversity. Yet, this context also presents an opportunity: the fixed relationship between the camera and the ground plane imposes additional perspective geometry constraints, enabling depth regression via vertical image positions of objects. However, this cue is highly susceptible to overfitting, thus we propose a novel canonical representation that maintains consistency across varied camera setups, effectively disentangling depth from specific parameters and enhancing generalization across datasets. We also propose a novel architecture that adaptively and probabilistically fuses depths estimated via object size and vertical image position cues. A comprehensive evaluation demonstrates the effectiveness of the proposed approach on five autonomous driving datasets, achieving accurate metric depth estimation for varying resolutions, aspect ratios and camera setups. Notably, we achieve comparable accuracy to existing zero-shot methods, despite training on a single dataset with a single-camera setup.

📄 PDF Abstract BibTeX arXiv:2412.06080

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous DrivingAutonomous VehiclesDepth EstimationMonocular Depth Estimation

Similar Papers 제목 키워드 기반

Metric3Dv2: A Versatile Monocular Geometric Foundation Model for Zero-shot Metric Depth and Surface Normal Estimation

2024-03-22 · Under review for Transaction 2024 4 · Mu Hu, Wei Yin, Chi Zhang, Zhipeng Cai 외

We introduce Metric3D v2, a geometric foundation model for zero-shot metric depth and surface normal estimation from a single image, which is crucial for metric 3D recovery. While depth and normal are geometrically relat…

Depth EstimationSurface Normal EstimationZero-shot Generalization

Can Language Understand Depth?

2022-07-03 · Renrui Zhang, Ziyao Zeng, Ziyu Guo, Yafeng Li

Besides image classification, Contrastive Language-Image Pre-training (CLIP) has accomplished extraordinary success for a wide range of vision tasks, including object-level and 3D space understanding. However, it's still…

Depth Estimationimage-classificationImage ClassificationMonocular Depth Estimation

Towards Zero-Shot Scale-Aware Monocular Depth Estimation

2023-06-29 · ICCV 2023 1 · Vitor Guizilini, Igor Vasiljevic, Dian Chen, Rares Ambrus 외

Monocular depth estimation is scale-ambiguous, and thus requires scale supervision to produce metric predictions. Even so, the resulting models will be geometry-specific, with learned scales that cannot be directly trans…

DecoderDepth EstimationMonocular Depth Estimation

ZipDepth: Bringing Lightweight Zero-Shot Monocular Depth Anywhere, on Any Device

2026-07-09 · Fabio Tosi, Luca Bartolomei, Matteo Poggi, Stefano Mattoccia arxiv

Monocular depth estimation has seen remarkable progress through foundation models achieving robust zero-shot generalization, yet their computational demands place them far beyond the reach of embedded and mobile platform…

Monocular Depth EstimationZero-shot GeneralizationKnowledge Distillation

Metric3D: Towards Zero-shot Metric 3D Prediction from A Single Image

2023-07-20 · ICCV 2023 1 · Wei Yin, Chi Zhang, Hao Chen, Zhipeng Cai 외

Reconstructing accurate 3D scenes from images is a long-standing vision task. Due to the ill-posedness of the single-image reconstruction problem, most well-established methods are built upon multi-view geometry. State-o…

Depth EstimationImage ReconstructionMonocular Depth EstimationZero-shot Generalization