CORAL: Colored structural representation for bi-modal place recognition
Place recognition is indispensable for a drift-free localization system. Due to the variations of the environment, place recognition using single-modality has limitations. In this paper, we propose a bi-modal place recognition method, which can extract a compound global descriptor from the two modalities, vision and LiDAR. Specifically, we first build the elevation image generated from 3D points as a structural representation. Then, we derive the correspondences between 3D points and image pixels that are further used in merging the pixel-wise visual features into the elevation map grids. In this way, we fuse the structural features and visual features in the consistent bird-eye view frame, yielding a semantic representation, namely CORAL. And the whole network is called CORAL-VLAD. Comparisons on the Oxford RobotCar show that CORAL-VLAD has superior performance against other state-of-the-art methods. We also demonstrate that our network can be generalized to other scenes and sensor configurations on cross-city datasets.
Code (0)
등록된 구현이 없습니다.
Tasks
Visual Place RecognitionMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Orthogonalized Multimodal Contrastive Learning with Asymmetric Masking for Structured Representations
Multimodal learning seeks to integrate information from heterogeneous sources, where signals may be shared across modalities, specific to individual modalities, or emerge only through their interaction. While self-superv…
Contrastive LearningUnderwater Soft Coral Detection: SCoralNet for Accurate and Efficient Annotation.
SCoralNet (based on Faster R-CNN) is a new underwater coral detection framework that has been proposed to automatically localize and identify distinct coral species in images, allowing for fast and detailed annotation. M…
2D Object DetectionColored Jones Polynomials and the Volume Conjecture
Using the vertex model approach for braid representations, we compute polynomials for spin-1 placed on hyperbolic knots up to 15 crossings. These polynomials are referred to as 3-colored Jones polynomials or adjoint Jone…
CoralBay: A Self-Supervised CT Foundation Model
Self-supervised learning has enabled large-scale pre-training on 2D natural images, producing general-purpose visual representations that transfer effectively across tasks. However, many medical imaging modalities, such …
Self-Supervised LearningRepresentation LearningCombining Photogrammetric Computer Vision and Semantic Segmentation for Fine-grained Understanding of Coral Reef Growth under Climate Change
Corals are the primary habitat-building life-form on reefs that support a quarter of the species in the ocean. A coral reef ecosystem usually consists of reefs, each of which is like a tall building in any city. These re…
Semantic Segmentation