paper-with-me

홈 › Papers

MS-Net: A Multi-modal Self-supervised Network for Fine-Grained Classification of Aircraft in SAR Images

2023-08-28 · Bingying Yue, Jianhao Li, Hao Shi, Yupei Wang, Honghu Zhong

Synthetic aperture radar (SAR) imaging technology is commonly used to provide 24-hour all-weather earth observation. However, it still has some drawbacks in SAR target classification, especially in fine-grained classification of aircraft: aircrafts in SAR images have large intra-class diversity and inter-class similarity; the number of effective samples is insufficient and it's hard to annotate. To address these issues, this article proposes a novel multi-modal self-supervised network (MS-Net) for fine-grained classification of aircraft. Firstly, in order to entirely exploit the potential of multi-modal information, a two-sided path feature extraction network (TSFE-N) is constructed to enhance the image feature of the target and obtain the domain knowledge feature of text mode. Secondly, a contrastive self-supervised learning (CSSL) framework is employed to effectively learn useful label-independent feature from unbalanced data, a similarity per-ception loss (SPloss) is proposed to avoid network overfitting. Finally, TSFE-N is used as the encoder of CSSL to obtain the classification results. Through a large number of experiments, our MS-Net can effectively reduce the difficulty of classifying similar types of aircrafts. In the case of no label, the proposed algorithm achieves an accuracy of 88.46% for 17 types of air-craft classification task, which has pioneering significance in the field of fine-grained classification of aircraft in SAR images.

📄 PDF Abstract BibTeX arXiv:2308.14613

Code (0)

등록된 구현이 없습니다.

Tasks

ClassificationEarth ObservationSelf-Supervised Learning

Similar Papers 제목 키워드 기반

Fine-grained Multi-Modal Self-Supervised Learning

2021-12-22 · Duo Wang, Salah Karout

Multi-Modal Self-Supervised Learning from videos has been shown to improve model's performance on various downstream tasks. However, such Self-Supervised pre-training requires large batch sizes and a large amount of comp…

Action RecognitionSelf-Supervised Learning

Self-Supervised Multimodal Learning: A Survey

2023-03-31 · Yongshuo Zong, Oisin Mac Aodha, Timothy Hospedales

Multimodal learning, which aims to understand and analyze information from multiple modalities, has achieved substantial progress in the supervised regime in recent years. However, the heavy dependence on data paired wit…

Machine TranslationSelf-Supervised LearningSurvey

Multi-Modal Domain Adaptation for Fine-Grained Action Recognition

2020-01-27 · CVPR 2020 6 · Jonathan Munro, Dima Damen

Fine-grained action recognition datasets exhibit environmental bias, where multiple video sequences are captured from a limited number of environments. Training a model in one environment and deploying in another results…

Action RecognitionDomain AdaptationFine-grained Action RecognitionOptical Flow Estimation+1

Self-Supervised Predictive Coding with Multimodal Fusion for Patient Deterioration Prediction in Fine-grained Time Resolution

2022-10-29 · Kwanhyung Lee, John Won, Heejung Hyun, Sangchul Hahn 외

Accurate time prediction of patients' critical events is crucial in urgent scenarios where timely decision-making is important. Though many studies have proposed automatic prediction methods using Electronic Health Recor…

Decision MakingFuture predictionPredictionTime Series Analysis

Towards Fine-Grained Adaptation of CLIP via a Self-Trained Alignment Score

2025-07-13 · Eman Ali, Sathira Silva, Chetan Arora, Muhammad Haris Khan arxiv

Vision-language models (VLMs) like CLIP excel in zero-shot learning by aligning image and text representations through contrastive pretraining. Existing approaches to unsupervised adaptation (UA) for fine-grained classif…

Zero-Shot Learning