paper-with-me

홈 › Papers

GSAlign: Geometric and Semantic Alignment Network for Aerial-Ground Person Re-Identification

2025-10-25 · Qiao Li, Jie Li, Yukang Zhang, Lei Tan, Jing Chen, Jiayi Ji arxiv

Aerial-Ground person re-identification (AG-ReID) is an emerging yet challenging task that aims to match pedestrian images captured from drastically different viewpoints, typically from unmanned aerial vehicles (UAVs) and ground-based surveillance cameras. The task poses significant challenges due to extreme viewpoint discrepancies, occlusions, and domain gaps between aerial and ground imagery. While prior works have made progress by learning cross-view representations, they remain limited in handling severe pose variations and spatial misalignment. To address these issues, we propose a Geometric and Semantic Alignment Network (GSAlign) tailored for AG-ReID. GSAlign introduces two key components to jointly tackle geometric distortion and semantic misalignment in aerial-ground matching: a Learnable Thin Plate Spline (LTPS) Module and a Dynamic Alignment Module (DAM). The LTPS module adaptively warps pedestrian features based on a set of learned keypoints, effectively compensating for geometric variations caused by extreme viewpoint changes. In parallel, the DAM estimates visibility-aware representation masks that highlight visible body regions at the semantic level, thereby alleviating the negative impact of occlusions and partial observations in cross-view correspondence. A comprehensive evaluation on CARGO with four matching protocols demonstrates the effectiveness of GSAlign, achieving significant improvements of +18.8\% in mAP and +16.8\% in Rank-1 accuracy over previous state-of-the-art methods on the aerial-ground setting.

📄 PDF Abstract BibTeX arXiv:2510.22268

Code (0)

등록된 구현이 없습니다.

Tasks

Person Re-Identification

Similar Papers 제목 키워드 기반

Cross-modal Fuzzy Alignment Network for Text-Aerial Person Retrieval and A Large-scale Benchmark

2026-03-21 · Yifei Deng, Chenglong Li, Yuyang Zhang, Guyue Hu 외 arxiv

Text-aerial person retrieval aims to identify targets in UAV-captured images from eyewitness descriptions, supporting intelligent transportation and public security applications. Compared to ground-view text--image perso…

Person RetrievalText Generation

SMDT: Cross-View Geo-Localization with Image Alignment and Transformer

2022-04-06 · IEEE International Conference on Multimedia and Expo 2022 2022 4 · Xiaoyang Tian, Jie Shao, Deqiang Ouyang, Anjie Zhu 외

The goal of cross-view geo-localization is to determine the location of a given ground image by matching with aerial images. However, existing methods ignore the variability of scenes, additional information and spatial …

geo-localizationSegmentationSemantic Segmentation

Seeing Where to Deploy: Metric RGB-Based Traversability Analysis for Aerial-to-Ground Hidden Space Inspection

2026-03-15 · Seoyoung Lee, Shaekh Mohammad Shithil, Durgakant Pushp, Lantao Liu 외 arxiv

Inspection of confined infrastructure such as culverts often requires accessing hidden spaces whose entrances are reachable primarily from elevated viewpoints. Aerial-ground cooperation enables a UAV to deploy a compact …

Semantic Segmentation

SegFly: A Dataset and 2D-3D-2D Paradigm for Aerial RGB-Thermal Semantic Segmentation at Scale

2026-03-18 · Markus Gross, Sai Bharadhwaj Matha, Rui Song, Viswanathan Muthuveerappan 외 arxiv

Semantic segmentation for uncrewed aerial vehicles (UAVs) is fundamental for aerial scene understanding, yet existing RGB and RGB-T datasets remain limited in scale, diversity, and annotation efficiency due to the high c…

Semantic SegmentationScene UnderstandingImage Registration

Semantic-aware Network for Aerial-to-Ground Image Synthesis

2023-08-14 · Jinhyun Jang, Taeyong Song, Kwanghoon Sohn

Aerial-to-ground image synthesis is an emerging and challenging problem that aims to synthesize a ground image from an aerial image. Due to the highly different layout and object representation between the aerial and gro…

Image Generation