paper-with-me

Papers

Geographical Knowledge-driven Representation Learning for Remote Sensing Images

2021-07-12 · Wenyuan Li, Keyan Chen, Hao Chen, Zhenwei Shi

The proliferation of remote sensing satellites has resulted in a massive amount of remote sensing images. However, due to human and material resource constraints, the vast majority of remote sensing images remain unlabeled. As a result, it cannot be applied to currently available deep learning methods. To fully utilize the remaining unlabeled images, we propose a Geographical Knowledge-driven Representation learning method for remote sensing images (GeoKR), improving network performance and reduce the demand for annotated data. The global land cover products and geographical location associated with each remote sensing image are regarded as geographical knowledge to provide supervision for representation learning and network pre-training. An efficient pre-training framework is proposed to eliminate the supervision noises caused by imaging times and resolutions difference between remote sensing images and geographical knowledge. A large scale pre-training dataset Levir-KR is proposed to support network pre-training. It contains 1,431,950 remote sensing images from Gaofen series satellites with various resolutions. Experimental results demonstrate that our proposed method outperforms ImageNet pre-training and self-supervised representation learning methods and significantly reduces the burden of data annotation on downstream tasks such as scene classification, semantic segmentation, object detection, and cloud / snow detection. It demonstrates that our proposed method can be used as a novel paradigm for pre-training neural networks. Codes will be available on https://github.com/flyakon/Geographical-Knowledge-driven-Representaion-Learning.

📄 PDF Abstract BibTeX arXiv:2107.05276

Code (1)

flyakon/Geographical-Knowledge-driven-Representaion-Learning 공식 구현 pytorch

Tasks

object-detectionObject DetectionRepresentation LearningScene ClassificationSemantic Segmentation

Similar Papers 제목 키워드 기반

Geospatial-Reasoning-Driven Vocabulary-Agnostic Remote Sensing Semantic Segmentation

2026-02-09 · Chufeng Zhou, Jian Wang, Xinyuan Liu, Xiaokang Zhang arxiv

Open-vocabulary semantic segmentation has become an important direction in remote sensing, as it enables recognition beyond predefined land-cover categories. However, existing methods mainly depend on passive visual-text…

Knowledge DistillationSemantic Segmentation

SFR-Net: Learning Scale-Frustum Representations for Ultra-Wide Area Remote Sensing Image Segmentation

2026-05-25 · Chuyu Zhong, Keyan Chen, Qinzhe Yang, Bowen Chen 외 arxiv

Pixel count and geographical coverage are two key characteristics of remote sensing images. Existing remote sensing image segmentation methods typically focus on images with either a small pixel count or a large pixel co…

Image Segmentation

Deep Learning Based Domain Adaptation Methods in Remote Sensing: A Comprehensive Survey

2025-10-17 · Shuchang Lyu, Qi Zhao, Zheng Zhou, Meng Li 외 arxiv

Domain adaptation is a crucial and increasingly important task in remote sensing, aiming to transfer knowledge from a source domain a differently distributed target domain. It has broad applications across various real-w…

Domain Adaptation

GeoMag: A Vision-Language Model for Pixel-level Fine-Grained Remote Sensing Image Parsing

2025-07-08 · Xianzhi Ma, Jianhui Li, Changhua Pei, Hao liu

The application of Vision-Language Models (VLMs) in remote sensing (RS) image understanding has achieved notable progress, demonstrating the basic ability to recognize and describe geographical entities. However, existin…

Language ModelingLanguage ModellingObject Recognition

GeoLLaVA: Efficient Fine-Tuned Vision-Language Models for Temporal Change Detection in Remote Sensing

2024-10-25 · Hosam Elgendy, Ahmed Sharshar, Ahmed Aboeitta, Yasser Ashraf 외

Detecting temporal changes in geographical landscapes is critical for applications like environmental monitoring and urban planning. While remote sensing data is abundant, existing vision-language models (VLMs) often fai…

Change Detection