paper-with-me

홈 › Papers

Satellite Image Based Cross-view Localization for Autonomous Vehicle

2022-07-27 · Shan Wang, Yanhao Zhang, Ankit Vora, Akhil Perincherry, Hongdong Li

Existing spatial localization techniques for autonomous vehicles mostly use a pre-built 3D-HD map, often constructed using a survey-grade 3D mapping vehicle, which is not only expensive but also laborious. This paper shows that by using an off-the-shelf high-definition satellite image as a ready-to-use map, we are able to achieve cross-view vehicle localization up to a satisfactory accuracy, providing a cheaper and more practical way for localization. While the utilization of satellite imagery for cross-view localization is an established concept, the conventional methodology focuses primarily on image retrieval. This paper introduces a novel approach to cross-view localization that departs from the conventional image retrieval method. Specifically, our method develops (1) a Geometric-align Feature Extractor (GaFE) that leverages measured 3D points to bridge the geometric gap between ground and overhead views, (2) a Pose Aware Branch (PAB) adopting a triplet loss to encourage pose-aware feature extraction, and (3) a Recursive Pose Refine Branch (RPRB) using the Levenberg-Marquardt (LM) algorithm to align the initial pose towards the true vehicle pose iteratively. Our method is validated on KITTI and Ford Multi-AV Seasonal datasets as ground view and Google Maps as the satellite view. The results demonstrate the superiority of our method in cross-view localization with median spatial and angular errors within $1$ meter and $1^\circ$, respectively.

📄 PDF Abstract BibTeX arXiv:2207.13506

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous VehiclesImage RetrievalRetrievalTriplet

Methods 이 논문이 사용한 방법론

AWARE We propose to theoretically and empirically examine the effect of incorporating weighting schemes into walk-aggregating GNNs. To this end, we propose a simple, interpretable, and…
Triplet Loss The goal of Triplet loss, in the context of Siamese Networks, is to maximize the joint probability among all score-pairs i.e. the product of all probabilities. By using its…
ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

Beyond Cross-view Image Retrieval: Highly Accurate Vehicle Localization Using Satellite Image

2022-04-10 · CVPR 2022 1 · Yujiao Shi, Hongdong Li

This paper addresses the problem of vehicle-mounted camera localization by matching a ground-level image with an overhead-view satellite map. Existing methods often treat this problem as cross-view image retrieval, and u…

Camera LocalizationImage RetrievalPose EstimationRetrieval

UAV-VisLoc: A Large-scale Dataset for UAV Visual Localization

2024-05-20 · Wenjia Xu, Yaxuan Yao, Jiaqi Cao, Zhiwei Wei 외

The application of unmanned aerial vehicles (UAV) has been widely extended recently. It is crucial to ensure accurate latitude and longitude coordinates for UAVs, especially when the global navigation satellite systems (…

Visual Localization

Trajectory-aware Cross-view Geo-localization with Sequential Observations

2026-07-16 · Tianyi Gao, Jiayu Lin, Danielle Beaulieu, Nathan Jacobs arxiv

Cross-view geo-localization matches ground-level observations against geo-tagged satellite imagery. Recent methods show that sequential queries such as video clips yield richer spatiotemporal cues than single images, yet…

CVLNet: Cross-View Semantic Correspondence Learning for Video-based Camera Localization

2022-08-07 · Yujiao Shi, Xin Yu, Shan Wang, Hongdong Li

This paper tackles the problem of Cross-view Video-based camera Localization (CVL). The task is to localize a query camera by leveraging information from its past observations, i.e., a continuous sequence of images obser…

Camera LocalizationImage-Based LocalizationSemantic correspondence

VIRD: View-Invariant Representation through Dual-Axis Transformation for Cross-View Pose Estimation

2026-03-13 · Juhye Park, Wooju Lee, Dasol Hong, Changki Sung 외 arxiv

Accurate global localization is critical for autonomous driving and robotics, but GNSS-based approaches often degrade due to occlusion and multipath effects. As an emerging alternative, cross-view pose estimation predict…

Autonomous DrivingPose Estimation