Oriented Feature Alignment for Fine-grained Object Recognition in High-Resolution Satellite Imagery
Oriented object detection in remote sensing images has made great progress in recent years. However, most of the current methods only focus on detecting targets, and cannot distinguish fine-grained objects well in complex scenes. In this technical report, we analyzed the key issues of fine-grained object recognition, and use an oriented feature alignment network (OFA-Net) to achieve high-performance fine-grained oriented object recognition in optical remote sensing images. OFA-Net achieves accurate object localization through a rotated bounding boxes refinement module. On this basis, the boundary-constrained rotation feature alignment module is applied to achieve local feature extraction, which is beneficial to fine-grained object classification. The single model of our method achieved mAP of 46.51\% in the GaoFen competition and won 3rd place in the ISPRS benchmark with the mAP of 43.73\%.
Code (0)
등록된 구현이 없습니다.
Tasks
Objectobject-detectionObject DetectionObject LocalizationObject RecognitionOriented Object DetectionSimilar Papers 제목 키워드 기반
TOAN: Target-Oriented Alignment Network for Fine-Grained Image Categorization with Few Labeled Samples
The challenges of high intra-class variance yet low inter-class fluctuations in fine-grained visual categorization are more severe with few labeled samples, \textit{i.e.,} Fine-Grained categorization problems under the F…
Fine-Grained Visual CategorizationImage CategorizationLearn Temporal Consistency For Robust Satellite Video Detector
Satellite video object detection (SVOD) for oriented and fine-grained objects plays an important role in satellite applications. Most existing SVOD methods only focus on one or a few coarse-grained categories of moving o…
Representation LearningVideo Object DetectionBenchmarking Large Vision-Language Models on Fine-Grained Image Tasks: From Evaluation to Diagnosis
Recent advancements in Large Vision-Language Models (LVLMs) have demonstrated remarkable multimodal perception and reasoning capabilities. While numerous benchmarks have evaluated LVLMs from holistic or task-specific per…
TTPA: Token-level Tool-use Preference Alignment Training Framework with Fine-grained Evaluation
Existing tool-learning methods usually rely on supervised fine-tuning, they often overlook fine-grained optimization of internal tool call details, leading to limitations in preference alignment and error discrimination.…
Spatial-Scale Aligned Network for Fine-Grained Recognition
Existing approaches for fine-grained visual recognition focus on learning marginal region-based representations while neglecting the spatial and scale misalignments, leading to inferior performance. In this paper, we pro…
Fine-Grained Visual Recognition