MSG-SR-Net: A Weakly Supervised Network Integrating Multi-Scale Generation and Super-Pixel Refinement for Building Extraction from High-Resolution Remotely Sensed Imageries
Weakly supervised semantic segmentation (WSSS) methods based on image-level labels can relieve the tedious pixel-level annotation burden, and these methods are mainly based on class activation maps (CAMs). However, it is challenging to generate high-quality CAMs for high-resolution remotely sensed imagery (HRSI). In this paper, we propose a WSSS method for building extraction from HRSI using image-level labels. The proposed method, termed as the MSG-SR-Net, integrates two novel modules, i.e., multi-scale generation (MSG) and super-pixel refinement (SR), to obtain high-quality CAMs so as to provide reliable pixel-level training samples for subsequent semantic segmentation steps. The MSG module is proposed to use global semantic information to guide the learning of multiple features across different levels, and then respectively to utilize multi-level features for generating multi-scale CAMs. This component can effectively suppress the interference of the class-irrelevant noise and strengthen the use of profitable information in multi-level features. The SR module is designed to take advantage of super-pixels to improve multi-scale CAMs in target integrity and details preserving. Extensive experiments on two public building datasets demonstrated that the proposed modules made the MSG-SR-Net obtain more integral and accurate CAMs for building extraction. Moreover, experimental results also showed the proposed method achieved excellent performance with over 67% in F1-score, and outperformed other weakly supervised methods in effectiveness and generalization ability.
Code (1)
Tasks
Semantic SegmentationWeakly supervised Semantic SegmentationWeakly-Supervised Semantic SegmentationSimilar Papers 제목 키워드 기반
DAWN: Domain-Adaptive Weakly Supervised Nuclei Segmentation via Cross-Task Interactions
Weakly supervised segmentation methods have gained significant attention due to their ability to reduce the reliance on costly pixel-level annotations during model training. However, the current weakly supervised nuclei …
Domain AdaptationPseudo LabelSegmentationWeakly supervised segmentationSelf-supervised Scale Equivariant Network for Weakly Supervised Semantic Segmentation
Weakly supervised semantic segmentation has attracted much research interest in recent years considering its advantage of low labeling cost. Most of the advanced algorithms follow the design principle that expands and co…
SegmentationSemantic SegmentationWeakly-supervised LearningWeakly supervised Semantic Segmentation+1DSNet: A Dual-Stream Framework for Weakly-Supervised Gigapixel Pathology Image Analysis
We present a novel weakly-supervised framework for classifying whole slide images (WSIs). WSIs, due to their gigapixel resolution, are commonly processed by patch-wise classification with patch-level labels. However, pat…
Classificationwhole slide imagesTraining ASR models by Generation of Contextual Information
Supervised ASR models have reached unprecedented levels of accuracy, thanks in part to ever-increasing amounts of labelled training data. However, in many applications and locales, only moderate amounts of data are avail…
Decoderspeech-recognitionSpeech RecognitionText Generation+1Curriculum Point Prompting for Weakly-Supervised Referring Image Segmentation
Referring image segmentation (RIS) aims to precisely segment referents in images through corresponding natural language expressions, yet relying on cost-intensive mask annotations. Weakly supervised RIS thus learns from …
Image SegmentationSegmentationSemantic Segmentation