paper-with-me

Papers

Slender Object Scene Segmentation in Remote Sensing Image Based on Learnable Morphological Skeleton with Segment Anything Model

2024-11-13 · Jun Xie, Wenxiao Li, Faqiang Wang, Liqiang Zhang, Zhengyang Hou, Jun Liu

Morphological methods play a crucial role in remote sensing image processing, due to their ability to capture and preserve small structural details. However, most of the existing deep learning models for semantic segmentation are based on the encoder-decoder architecture including U-net and Segment Anything Model (SAM), where the downsampling process tends to discard fine details. In this paper, we propose a new approach that integrates learnable morphological skeleton prior into deep neural networks using the variational method. To address the difficulty in backpropagation in neural networks caused by the non-differentiability presented in classical morphological operations, we provide a smooth representation of the morphological skeleton and design a variational segmentation model integrating morphological skeleton prior by employing operator splitting and dual methods. Then, we integrate this model into the network architecture of SAM, which is achieved by adding a token to mask decoder and modifying the final sigmoid layer, ensuring the final segmentation results preserve the skeleton structure as much as possible. Experimental results on remote sensing datasets, including buildings and roads, demonstrate that our method outperforms the original SAM on slender object segmentation and exhibits better generalization capability.

📄 PDF Abstract BibTeX arXiv:2411.08592

Code (0)

등록된 구현이 없습니다.

Tasks

DecoderScene SegmentationSegmentationSemantic Segmentation

Methods 이 논문이 사용한 방법론

Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…
ReLU How Do I Communicate to Expedia? How Do I Communicate to Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Live Support & Special Travel…
Concatenated Skip Connection A Concatenated Skip Connection is a type of skip connection that seeks to reuse features by concatenating them to new layers, allowing more information to be retained from…
Max Pooling Max Pooling is a pooling operation that calculates the maximum value for patches of a feature map, and uses it to create a downsampled (pooled) feature map. It is usually…
U-Net 설명 없음
SAM 설명 없음

Similar Papers 제목 키워드 기반

SACANet: scene-aware class attention network for semantic segmentation of remote sensing images

2023-04-22 · Xiaowen Ma, Rui Che, Tingfeng Hong, Mengting Ma 외

Spatial attention mechanism has been widely used in semantic segmentation of remote sensing images given its capability to model long-range dependencies. Many methods adopting spatial attention mechanism aggregate contex…

Semantic Segmentation

SCORE: Scene Context Matters in Open-Vocabulary Remote Sensing Instance Segmentation

2025-07-17 · Shiqi Huang, Shuting He, Huaiyuan Qin, Bihan Wen

Most existing remote sensing instance segmentation approaches are designed for close-vocabulary prediction, limiting their ability to recognize novel categories or generalize across datasets. This restricts their applica…

Earth ObservationInstance SegmentationSegmentationSemantic Segmentation

Class Attention Network for Semantic Segmentation of Remote Sensing Images

2020-12-31 · Zhibo Rao, Mingyi He, Yuchao Dai

Semantic segmentation in remote sensing images is beneficial to detect objects and understand the scene in earth observation. However, classical networks always failed to obtain an accuracy segmentation map in remote sen…

Earth ObservationScene ParsingSegmentationSemantic Segmentation

PKINet-v2: Towards Powerful and Efficient Poly-Kernel Remote Sensing Object Detection

2026-03-17 · Xinhao Cai, Liulei Li, Gensheng Pei, Zeren Sun 외 arxiv

Object detection in remote sensing images (RSIs) is challenged by the coexistence of geometric and spatial complexity: targets may appear with diverse aspect ratios, while spanning a wide range of object sizes under vari…

Object Detection

Generic Knowledge Boosted Pre-training For Remote Sensing Images

2024-01-09 · Ziyue Huang, Mingming Zhang, Yuan Gong, Qingjie Liu 외

Deep learning models are essential for scene classification, change detection, land cover segmentation, and other remote sensing image understanding tasks. Most backbones of existing remote sensing deep learning models a…

Change DetectionDeep LearningGeneral Knowledgeobject-detection+3