paper-with-me

홈 › Papers

Boosting Multi-View Stereo with Depth Foundation Model in the Absence of Real-World Labels

2025-04-16 · Jie Zhu, Bo Peng, Zhe Zhang, Bingzheng Liu, Jianjun Lei

Learning-based Multi-View Stereo (MVS) methods have made remarkable progress in recent years. However, how to effectively train the network without using real-world labels remains a challenging problem. In this paper, driven by the recent advancements of vision foundation models, a novel method termed DFM-MVS, is proposed to leverage the depth foundation model to generate the effective depth prior, so as to boost MVS in the absence of real-world labels. Specifically, a depth prior-based pseudo-supervised training mechanism is developed to simulate realistic stereo correspondences using the generated depth prior, thereby constructing effective supervision for the MVS network. Besides, a depth prior-guided error correction strategy is presented to leverage the depth prior as guidance to mitigate the error propagation problem inherent in the widely-used coarse-to-fine network structure. Experimental results on DTU and Tanks & Temples datasets demonstrate that the proposed DFM-MVS significantly outperforms existing MVS methods without using real-world labels.

📄 PDF Abstract BibTeX arXiv:2504.11845

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Boosting Omnidirectional Stereo Matching with a Pre-trained Depth Foundation Model

2025-03-30 · Jannik Endres, Oliver Hahn, Charles Corbière, Simone Schaub-Meyer 외

Omnidirectional depth perception is essential for mobile robotics applications that require scene understanding across a full 360{\deg} field of view. Camera-based setups offer a cost-effective option by using stereo dep…

Depth EstimationMonocular Depth EstimationOmnnidirectional Stereo Depth EstimationScene Understanding+2

Boosting Zero-shot Stereo Matching using Large-scale Mixed Images Sources in the Real World

2025-05-13 · Yuran Wang, Yingping Liang, Ying Fu

Stereo matching methods rely on dense pixel-wise ground truth labels, which are laborious to obtain, especially for real-world datasets. The scarcity of labeled data and domain gaps between synthetic and real-world image…

Depth EstimationMonocular Depth EstimationStereo Matching

Geometric Distillation from Rectified Stereo: Leveraging Epipolar Cues for Monocular Depth

2026-07-17 · Jung-Hee Kim, Xiaoming Liu arxiv

Monocular depth foundation models have demonstrated remarkable generalization capabilities across diverse environments. However, they continue to struggle with metric depth estimation in diverse environments. This limita…

Depth Estimation

BEVStereo++: Accurate Depth Estimation in Multi-view 3D Object Detection via Dynamic Temporal Stereo

2023-04-09 · Yinhao Li, Jinrong Yang, Jianjian Sun, Han Bao 외

Bounded by the inherent ambiguity of depth perception, contemporary multi-view 3D object detection methods fall into the performance bottleneck. Intuitively, leveraging temporal multi-view stereo (MVS) technology is the …

3D Object DetectionDepth EstimationMotion Compensationobject-detection+1

Multi view stereo with semantic priors

2020-07-05 · Elisavet Konstantina Stathopoulou, Fabio Remondino

Patch-based stereo is nowadays a commonly used image-based technique for dense 3D reconstruction in large scale multi-view applications. The typical steps of such a pipeline can be summarized in stereo pair selection, de…

3D Reconstruction