paper-with-me

홈 › Papers

Jointly Optimized Global-Local Visual Localization of UAVs

2023-10-12 · Haoling Li, Jiuniu Wang, Zhiwei Wei, Wenjia Xu

Navigation and localization of UAVs present a challenge when global navigation satellite systems (GNSS) are disrupted and unreliable. Traditional techniques, such as simultaneous localization and mapping (SLAM) and visual odometry (VO), exhibit certain limitations in furnishing absolute coordinates and mitigating error accumulation. Existing visual localization methods achieve autonomous visual localization without error accumulation by matching with ortho satellite images. However, doing so cannot guarantee real-time performance due to the complex matching process. To address these challenges, we propose a novel Global-Local Visual Localization (GLVL) network. Our GLVL network is a two-stage visual localization approach, combining a large-scale retrieval module that finds similar regions with the UAV flight scene, and a fine-grained matching module that localizes the precise UAV coordinate, enabling real-time and precise localization. The training process is jointly optimized in an end-to-end manner to further enhance the model capability. Experiments on six UAV flight scenes encompassing both texture-rich and texture-sparse regions demonstrate the ability of our model to achieve the real-time precise localization requirements of UAVs. Particularly, our method achieves a localization error of only 2.39 meters in 0.48 seconds in a village scene with sparse texture features.

📄 PDF Abstract BibTeX arXiv:2310.08082

Code (0)

등록된 구현이 없습니다.

Tasks

RetrievalSimultaneous Localization and MappingVisual LocalizationVisual Odometry

Similar Papers 제목 키워드 기반

Deep Global-Relative Networks for End-to-End 6-DoF Visual Localization and Odometry

2018-12-19 · Yimin Lin, Zhaoxiang Liu, Jianfeng Huang, Chaopeng Wang 외

Although a wide variety of deep neural networks for robust Visual Odometry (VO) can be found in the literature, they are still unable to solve the drift problem in long-term robot navigation. Thus, this paper aims to pro…

Autonomous NavigationPose EstimationRobot NavigationVisual Localization+1

Learning to Locate Visual Answer in Video Corpus Using Question

2022-10-11 · Bin Li, Yixuan Weng, Bin Sun, Shutao Li

We introduce a new task, named video corpus visual answer localization (VCVAL), which aims to locate the visual answer in a large collection of untrimmed instructional videos using a natural language question. This task …

Contrastive LearningLanguage ModellingRetrievalVideo Retrieval

CT-VIR: Continuous-Time Visual-Inertial-Ranging Fusion for Indoor Localization with Sparse Anchors

2026-04-16 · Yu-An Liu, Li Zhang arxiv

Visual-inertial odometry (VIO) is widely used for mobile robot localization, but its long-term accuracy degrades without global constraints. Incorporating ranging sensors such as ultra-wideband (UWB) can mitigate drift; …

Dual-modality seq2seq network for audio-visual event localization

2019-02-20 · Yan-Bo Lin, Yu-Jhe Li, Yu-Chiang Frank Wang

Audio-visual event localization requires one to identify theevent which is both visible and audible in a video (eitherat a frame or video level). To address this task, we pro-pose a deep neural network named Audio-Visual…

audio-visual event localization

Enriching Local and Global Contexts for Temporal Action Localization

2021-07-27 · ICCV 2021 10 · Zixin Zhu, Wei Tang, Le Wang, Nanning Zheng 외

Effectively tackling the problem of temporal action localization (TAL) necessitates a visual representation that jointly pursues two confounding goals, i.e., fine-grained discrimination for temporal localization and suff…

Action ClassificationAction LocalizationRetrievalTemporal Action Localization+1