IBD-SLAM: Learning Image-Based Depth Fusion for Generalizable SLAM
In this paper we address the challenging problem of visual SLAM with neural scene representations. Recently neural scene representations have shown promise for SLAM to produce dense 3D scene reconstruction with high quality. However existing methods require scene-specific optimization leading to time-consuming mapping processes for each individual scene. To overcome this limitation we propose IBD-SLAM an Image-Based Depth fusion framework for generalizable SLAM. In particular we adopt a Neural Radiance Field (NeRF) for scene representation. Inspired by multi-view image-based rendering instead of learning a fixed-grid scene representation we propose to learn an image-based depth fusion model that fuses depth maps of multiple reference views into a xyz-map representation. Once trained this model can be applied to new uncalibrated monocular RGBD videos of unseen scenes without the need for retraining and reconstructs full 3D scenes efficiently with a light-weight pose optimization procedure. We thoroughly evaluate IBD-SLAM on public visual SLAM benchmarks outperforming the previous state-of-the-art while being 10x faster in the mapping stage. Project page: https://visual-ai.github.io/ibd-slam.
Code (0)
등록된 구현이 없습니다.
Tasks
3D Scene ReconstructionNeRFSimilar Papers 제목 키워드 기반
CodeMapping: Real-Time Dense Mapping for Sparse SLAM using Compact Scene Representations
We propose a novel dense mapping framework for sparse visual SLAM systems which leverages a compact scene representation. State-of-the-art sparse visual SLAM systems provide accurate and reliable estimates of the camera …
3D ReconstructionDepth EstimationScene UnderstandingA Front-End for Dense Monocular SLAM using a Learned Outlier Mask Prior
Recent achievements in depth prediction from a single RGB image have powered the new research area of combining convolutional neural networks (CNNs) with classical simultaneous localization and mapping (SLAM) algorithms.…
Depth EstimationDepth PredictionPredictionSimultaneous Localization and MappingCNN-SLAM: Real-time dense monocular SLAM with learned depth prediction
Given the recent advances in depth prediction from Convolutional Neural Networks (CNNs), this paper investigates how predicted depth maps from a deep neural network can be deployed for accurate and dense monocular recons…
Depth EstimationDepth PredictionMonocular ReconstructionPredictionProbabilistic Volumetric Fusion for Dense Monocular SLAM
We present a novel method to reconstruct 3D scenes from images by leveraging deep dense monocular SLAM and fast uncertainty propagation. The proposed approach is able to 3D reconstruct scenes densely, accurately, and in …
DeepRelativeFusion: Dense Monocular SLAM using Single-Image Relative Depth Prediction
In this paper, we propose a dense monocular SLAM system, named DeepRelativeFusion, that is capable to recover a globally consistent 3D structure. To this end, we use a visual SLAM algorithm to reliably recover the camera…
Depth EstimationDepth PredictionSimultaneous Localization and MappingVisual Navigation