paper-with-me

Papers

CMSG Cross-Media Semantic-Graph Feature Matching Algorithm for Autonomous Vehicle Relocalization

2023-05-15 · Shuhang Tan, Hengyu Liu, Zhiling Wang

Relocalization is the basis of map-based localization algorithms. Camera and LiDAR map-based methods are pervasive since their robustness under different scenarios. Generally, mapping and localization using the same sensor have better accuracy since matching features between the same type of data is easier. However, due to the camera's lack of 3D information and the high cost of LiDAR, cross-media methods are developing, which combined live image data and Lidar map. Although matching features between different media is challenging, we believe cross-media is the tendency for AV relocalization since its low cost and accuracy can be comparable to the same-sensor-based methods. In this paper, we propose CMSG, a novel cross-media algorithm for AV relocalization tasks. Semantic features are utilized for better interpretation the correlation between point clouds and image features. What's more, abstracted semantic graph nodes are introduced, and a graph network architecture is integrated to better extract the similarity of semantic features. Validation experiments are conducted on the KITTI odometry dataset. Our results show that CMSG can have comparable or even better accuracy compared to current single-sensor-based methods at a speed of 25 FPS on NVIDIA 1080 Ti GPU.

📄 PDF Abstract BibTeX arXiv:2305.08318

Code (0)

등록된 구현이 없습니다.

Tasks

GPU

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…

Similar Papers 제목 키워드 기반

How to Describe Images in a More Funny Way? Towards a Modular Approach to Cross-Modal Sarcasm Generation

2022-11-20 · Jie Ruan, Yue Wu, Xiaojun Wan, Yuesheng Zhu

Sarcasm generation has been investigated in previous studies by considering it as a text-to-text generation problem, i.e., generating a sarcastic sentence for an input sentence. In this paper, we study a new problem of c…

DescriptiveSentenceText Generation

Semantic Structure Enhanced Contrastive Adversarial Hash Network for Cross-media Representation Learning

2022-10-22 · ACM Multimedia 2022 10 · Meiyu Liang, Junping Du, Xiaowen Cao, Yang Yu 외

Deep cross-media hashing technology provides an efficient cross-media representation learning solution for cross-media search. However, the existing methods do not consider both fine-grained semantic features and semanti…

Representation Learning

Less is More: Information Bottleneck Denoised Multimedia Recommendation

2025-01-21 · Yonghui Yang, Le Wu, Zhuangzhuang He, Zhengwei Wu 외

Empowered by semantic-rich content information, multimedia recommendation has emerged as a potent personalized technique. Current endeavors center around harnessing multimedia content to refine item representation or unc…

Multimedia recommendation

Scientific and Technological Information Oriented Semantics-adversarial and Media-adversarial Cross-media Retrieval

2022-03-16 · Ang Li, Junping Du, Feifei Kou, Zhe Xue 외

Cross-media retrieval of scientific and technological information is one of the important tasks in the cross-media study. Cross-media scientific and technological information retrieval obtain target information from mass…

Information RetrievalRetrievalSemantic SimilaritySemantic Textual Similarity

Multi-Task Semantic Communication With Graph Attention-Based Feature Correlation Extraction

2025-01-02 · Xi Yu, Tiejun Lv, Weicai Li, Wei Ni 외

Multi-task semantic communication can serve multiple learning tasks using a shared encoder model. Existing models have overlooked the intricate relationships between features extracted during an encoding process of tasks…

Feature CorrelationGraph AttentionSemantic Communication