paper-with-me

홈 › Papers

Leveraging BEV Representation for 360-degree Visual Place Recognition

2023-05-23 · Xuecheng Xu, Yanmei Jiao, Sha Lu, Xiaqing Ding, Rong Xiong, Yue Wang

This paper investigates the advantages of using Bird's Eye View (BEV) representation in 360-degree visual place recognition (VPR). We propose a novel network architecture that utilizes the BEV representation in feature extraction, feature aggregation, and vision-LiDAR fusion, which bridges visual cues and spatial awareness. Our method extracts image features using standard convolutional networks and combines the features according to pre-defined 3D grid spatial points. To alleviate the mechanical and time misalignments between cameras, we further introduce deformable attention to learn the compensation. Upon the BEV feature representation, we then employ the polar transform and the Discrete Fourier transform for aggregation, which is shown to be rotation-invariant. In addition, the image and point cloud cues can be easily stated in the same coordinates, which benefits sensor fusion for place recognition. The proposed BEV-based method is evaluated in ablation and comparative studies on two datasets, including on-the-road and off-the-road scenarios. The experimental results verify the hypothesis that BEV can benefit VPR by its superior performance compared to baseline methods. To the best of our knowledge, this is the first trial of employing BEV representation in this task.

📄 PDF Abstract BibTeX arXiv:2305.13814

Code (1)

maverickpeter/vdisco 공식 구현 pytorch

Tasks

Sensor FusionVisual Place Recognition

Similar Papers 제목 키워드 기반

MSSPlace: Multi-Sensor Place Recognition with Visual and Text Semantics

2024-07-22 · Alexander Melekhin, Dmitry Yudin, Ilia Petryashin, Vitaly Bezuglyj

Place recognition is a challenging task in computer vision, crucial for enabling autonomous vehicles and robots to navigate previously visited environments. While significant progress has been made in learnable multimoda…

Autonomous VehiclesNavigateSemantic Segmentation

Improving Multimodal Speech Recognition by Data Augmentation and Speech Representations

2021-11-16 · ACL ARR November 2021 11 · Anonymous

Multimodal speech recognition aims to improve the performance of automatic speech recognition (ASR) systems by leveraging additional visual information that is usually associated to the audio input. While previous approa…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Data Augmentationspeech-recognition+1

Self-Supervised Visual Place Recognition Learning in Mobile Robots

2019-05-11 · Sudeep Pillai, John Leonard

Place recognition is a critical component in robot navigation that enables it to re-establish previously visited locations, and simultaneously use this information to correct the drift incurred in its dead-reckoned estim…

Metric LearningRobot NavigationVisual Place Recognition

Are State-of-the-art Visual Place Recognition Techniques any Good for Aerial Robotics?

2019-04-16 · Mubariz Zaffar, Ahmad Khaliq, Shoaib Ehsan, Michael Milford 외

Visual Place Recognition (VPR) has seen significant advances at the frontiers of matching performance and computational superiority over the past few years. However, these evaluations are performed for ground-based mobil…

Visual Place Recognition

Place recognition: An Overview of Vision Perspective

2017-06-17 · Zhiqiang Zeng, Jian Zhang, Xiaodong Wang, Yuming Chen 외

Place recognition is one of the most fundamental topics in computer vision and robotics communities, where the task is to accurately and efficiently recognize the location of a given query image. Despite years of wisdom …

image-classificationImage ClassificationImage DescriptionImage Retrieval+4