paper-with-me

홈 › Papers

InterKey: Cross-modal Intersection Keypoints for Global Localization on OpenStreetMap

2025-09-17 · Nguyen Hoang Khoi Tran, Julie Stephany Berrio, Mao Shan, Stewart Worrall arxiv

Reliable global localization is critical for autonomous vehicles, especially in environments where GNSS is degraded or unavailable, such as urban canyons and tunnels. Although high-definition (HD) maps provide accurate priors, the cost of data collection, map construction, and maintenance limits scalability. OpenStreetMap (OSM) offers a free and globally available alternative, but its coarse abstraction poses challenges for matching with sensor data. We propose InterKey, a cross-modal framework that leverages road intersections as distinctive landmarks for global localization. Our method constructs compact binary descriptors by jointly encoding road and building imprints from point clouds and OSM. To bridge modality gaps, we introduce discrepancy mitigation, orientation determination, and area-equalized sampling strategies, enabling robust cross-modal matching. Experiments on the KITTI dataset demonstrate that InterKey achieves state-of-the-art accuracy, outperforming recent baselines by a large margin. The framework generalizes to sensors that can produce dense structural point clouds, offering a scalable and cost-effective solution for robust vehicle localization.

📄 PDF Abstract BibTeX arXiv:2509.13857

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous VehiclesPoint Clouds

Similar Papers 제목 키워드 기반

Pedestrian Crossing Action Recognition and Trajectory Prediction with 3D Human Keypoints

2023-06-01 · Jiachen Li, Xinwei Shi, Feiyu Chen, Jonathan Stroud 외

Accurate understanding and prediction of human behaviors are critical prerequisites for autonomous vehicles, especially in highly dynamic and interactive scenarios such as intersections in dense urban areas. In this work…

Action RecognitionAutonomous VehiclesContrastive LearningMulti-Task Learning+1

Cross-Modal Information-Guided Network using Contrastive Learning for Point Cloud Registration

2023-11-02 · Yifan Xie, Jihua Zhu, Shiqi Li, Pengcheng Shi

The majority of point cloud registration methods currently rely on extracting features from points. However, these methods are limited by their dependence on information obtained from a single modality of points, which c…

Contrastive LearningPoint Cloud Registration

Pose2Instance: Harnessing Keypoints for Person Instance Segmentation

2017-04-04 · Subarna Tripathi, Maxwell Collins, Matthew Brown, Serge Belongie

Human keypoints are a well-studied representation of people.We explore how to use keypoint models to improve instance-level person segmentation. The main idea is to harness the notion of a distance transform of oracle pr…

Instance SegmentationSegmentationSemantic Segmentation

MD-Net: Multi-Detector for Local Feature Extraction

2022-08-10 · Emanuele Santellani, Christian Sormann, Mattia Rossi, Andreas Kuhn 외

Establishing a sparse set of keypoint correspon dences between images is a fundamental task in many computer vision pipelines. Often, this translates into a computationally expensive nearest neighbor search, where every …

3D Reconstruction

Deep Fusion Transformer Network with Weighted Vector-Wise Keypoints Voting for Robust 6D Object Pose Estimation

2023-08-10 · ICCV 2023 1 · Jun Zhou, Kai Chen, Linlin Xu, Qi Dou 외

One critical challenge in 6D object pose estimation from a single RGBD image is efficient integration of two different modalities, i.e., color and depth. In this work, we tackle this problem by a novel Deep Fusion Transf…

6D Pose Estimation using RGBglobal-optimizationPose EstimationSemantic Similarity+1