paper-with-me

홈 › Papers

FoundationSLAM: Unleashing the Power of Depth Foundation Models for End-to-End Dense Visual SLAM

2025-12-31 · Yuchen Wu, Jiahe Li, Fabio Tosi, Matteo Poggi, Jin Zheng, Xiao Bai arxiv

We present FoundationSLAM, a learning-based monocular dense SLAM system that addresses the absence of geometric consistency in previous flow-based approaches for accurate and robust tracking and mapping. Our core idea is to bridge flow estimation with geometric reasoning by leveraging the guidance from foundation depth models. To this end, we first develop a Hybrid Flow Network that produces geometry-aware correspondences, enabling consistent depth and pose inference across diverse keyframes. To enforce global consistency, we propose a Bi-Consistent Bundle Adjustment Layer that jointly optimizes keyframe pose and depth under multi-view constraints. Furthermore, we introduce a Reliability-Aware Refinement mechanism that dynamically adapts the flow update process by distinguishing between reliable and uncertain regions, forming a closed feedback loop between matching and optimization. Extensive experiments demonstrate that FoundationSLAM achieves superior trajectory accuracy and dense reconstruction quality across multiple challenging datasets, while running in real-time at 18 FPS, demonstrating strong generalization to various scenarios and practical applicability of our method.

📄 PDF Abstract BibTeX arXiv:2512.25008

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Prompting Depth Anything for 4K Resolution Accurate Metric Depth Estimation

2024-12-18 · CVPR 2025 1 · Haotong Lin, Sida Peng, Jingxiao Chen, Songyou Peng 외

Prompts play a critical role in unleashing the power of language and vision foundation models for specific tasks. For the first time, we introduce prompting into depth foundation models, creating a new paradigm for metri…

3D Reconstruction4kDecoderDepth Estimation+1

Depth Anything: Unleashing the Power of Large-Scale Unlabeled Data

2024-01-19 · CVPR 2024 1 · Lihe Yang, Bingyi Kang, Zilong Huang, Xiaogang Xu 외

This work presents Depth Anything, a highly practical solution for robust monocular depth estimation. Without pursuing novel technical modules, we aim to build a simple yet powerful foundation model dealing with any imag…

Data AugmentationDepth EstimationMonocular Depth EstimationSemantic Segmentation

Last-Layer-Centric Feature Recombination: Unleashing 3D Geometric Knowledge in DINOv3 for Monocular Depth Estimation

2026-04-29 · Gongshu Wang, Zhirui Wang, Kan Yang arxiv

Monocular depth estimation (MDE) is a fundamental yet inherently ill-posed task. Recent vision foundation models (VFMs), particularly DINO-based transformers, have significantly improved accuracy and generalization for d…

Monocular Depth Estimation

Distilling Monocular Foundation Model for Fine-grained Depth Completion

2025-01-01 · CVPR 2025 1 · Yingping Liang, Yutao Hu, Wenqi Shao, Ying Fu

Depth completion involves predicting dense depth maps from sparse LiDAR inputs, a critical task for applications such as autonomous driving and robotics. However, sparse depth annotations from sensors limit the avail…

Autonomous DrivingDepth CompletionDepth EstimationKnowledge Distillation+1

Unleashing the Power of LLMs in Dense Retrieval with Query Likelihood Modeling

2025-04-07 · Hengran Zhang, Keping Bi, Jiafeng Guo, Xiaojie Sun 외

Dense retrieval is a crucial task in Information Retrieval (IR) and is the foundation for downstream tasks such as re-ranking. Recently, large language models (LLMs) have shown compelling semantic understanding capabilit…

Information RetrievalLanguage ModelingLanguage ModellingRe-Ranking+2