paper-with-me

홈 › Papers

CoSAM: Self-Correcting SAM for Domain Generalization in 2D Medical Image Segmentation

2024-11-15 · Yihang Fu, Ziyang Chen, Yiwen Ye, Xingliang Lei, Zhisong Wang, Yong Xia

Medical images often exhibit distribution shifts due to variations in imaging protocols and scanners across different medical centers. Domain Generalization (DG) methods aim to train models on source domains that can generalize to unseen target domains. Recently, the segment anything model (SAM) has demonstrated strong generalization capabilities due to its prompt-based design, and has gained significant attention in image segmentation tasks. Existing SAM-based approaches attempt to address the need for manual prompts by introducing prompt generators that automatically generate these prompts. However, we argue that auto-generated prompts may not be sufficiently accurate under distribution shifts, potentially leading to incorrect predictions that still require manual verification and correction by clinicians. To address this challenge, we propose a method for 2D medical image segmentation called Self-Correcting SAM (CoSAM). Our approach begins by generating coarse masks using SAM in a prompt-free manner, providing prior prompts for the subsequent stages, and eliminating the need for prompt generators. To automatically refine these coarse masks, we introduce a generalized error decoder that simulates the correction process typically performed by clinicians. Furthermore, we generate diverse prompts as feedback based on the corrected masks, which are used to iteratively refine the predictions within a self-correcting loop, enhancing the generalization performance of our model. Extensive experiments on two medical image segmentation benchmarks across multiple scenarios demonstrate the superiority of CoSAM over state-of-the-art SAM-based methods.

📄 PDF Abstract BibTeX arXiv:2411.10136

Code (0)

등록된 구현이 없습니다.

Tasks

Domain GeneralizationImage SegmentationMedical Image SegmentationSemantic Segmentation

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음
SAM 설명 없음

Similar Papers 제목 키워드 기반

MedicoSAM: Towards foundation models for medical image segmentation

2025-01-20 · Anwai Archit, Luca Freckmann, Constantin Pape

Medical image segmentation is an important analysis task in clinical practice and research. Deep learning has massively advanced the field, but current approaches are mostly based on models trained for a specific task. T…

Image SegmentationInteractive SegmentationMedical Image SegmentationSegmentation+2

ANN-assisted CoSaMP Algorithm for Linear Electromagnetic Imaging of Spatially Sparse Domains

2019-11-15 · Ali I. Sandhu, Salman A. Shaukat, Abdulla Desmal, Hakan Bagci

Greedy pursuit algorithms (GPAs) are widely used to reconstruct sparse signals. Even though many electromagnetic (EM) inverse scattering problems are solved on sparse investigation domains, GPAs have rarely been used for…

Co-Correcting: Noise-tolerant Medical Image Classification via mutual Label Correction

2021-09-11 · Jiarun Liu, Ruirui Li, Chuan Sun

With the development of deep learning, medical image classification has been significantly improved. However, deep learning requires massive data with labels. While labeling the samples by human experts is expensive and …

ClassificationDeep Learningimage-classificationImage Classification+2

Continual Learning for Segment Anything Model Adaptation

2024-12-09 · Jinglong Yang, Yichen Wu, Jun Cen, Wenjian Huang 외

Although the current different types of SAM adaptation methods have achieved promising performance for various downstream tasks, such as prompt-based ones and adapter-based ones, most of them belong to the one-step adapt…

Continual Learningmodel

PicoSAM3: Real-Time In-Sensor Region-of-Interest Segmentation

2026-03-12 · Pietro Bonazzi, Nicola Farronato, Stefan Zihlmann, Haotong Qin 외 arxiv

Real-time, on-device segmentation is critical for latency-sensitive and privacy-aware applications such as smart glasses and Internet-of-Things devices. We introduce PicoSAM3, a lightweight promptable visual segmentation…

Knowledge Distillation