paper-with-me

홈 › Papers

Multi-scale Contrastive Adaptor Learning for Segmenting Anything in Underperformed Scenes

2024-08-12 · Ke Zhou, Zhongwei Qiu, Dongmei Fu

Foundational vision models, such as the Segment Anything Model (SAM), have achieved significant breakthroughs through extensive pre-training on large-scale visual datasets. Despite their general success, these models may fall short in specialized tasks with limited data, and fine-tuning such large-scale models is often not feasible. Current strategies involve incorporating adaptors into the pre-trained SAM to facilitate downstream task performance with minimal model adjustment. However, these strategies can be hampered by suboptimal learning approaches for the adaptors. In this paper, we introduce a novel Multi-scale Contrastive Adaptor learning method named MCA-SAM, which enhances adaptor performance through a meticulously designed contrastive learning framework at both token and sample levels. Our Token-level Contrastive adaptor (TC-adaptor) focuses on refining local representations by improving the discriminability of patch tokens, while the Sample-level Contrastive adaptor (SC-adaptor) amplifies global understanding across different samples. Together, these adaptors synergistically enhance feature comparison within and across samples, bolstering the model's representational strength and its ability to adapt to new tasks. Empirical results demonstrate that MCA-SAM sets new benchmarks, outperforming existing methods in three challenging domains: camouflage object detection, shadow segmentation, and polyp segmentation. Specifically, MCA-SAM exhibits substantial relative performance enhancements, achieving a 20.0% improvement in MAE on the COD10K dataset, a 6.0% improvement in MAE on the CAMO dataset, a 15.4% improvement in BER on the ISTD dataset, and a 7.9% improvement in mDice on the Kvasir-SEG dataset.

📄 PDF Abstract BibTeX arXiv:2408.05936

Code (0)

등록된 구현이 없습니다.

Tasks

Contrastive Learningobject-detectionObject Detection

Methods 이 논문이 사용한 방법론

SAM 설명 없음
Contrastive Learning 설명 없음
MAE 설명 없음

Similar Papers 제목 키워드 기반

Global-Local Medical SAM Adaptor Based on Full Adaption

2024-09-26 · Meng Wang, Yarong Feng, Yongwei Tang, Tian Zhang 외

Emerging of visual language models, such as the segment anything model (SAM), have made great breakthroughs in the field of universal semantic segmentation and significantly aid the improvements of medical image segmenta…

Image SegmentationMedical Image SegmentationSegmentationSemantic Segmentation

WeakMedSAM: Weakly-Supervised Medical Image Segmentation via SAM with Sub-Class Exploration and Prompt Affinity Mining

2025-03-06 · Haoran Wang, Lian Huai, Wenbin Li, Lei Qi 외

We have witnessed remarkable progress in foundation models in vision tasks. Currently, several recent works have utilized the segmenting anything model (SAM) to boost the segmentation performance in medical images, where…

Image SegmentationMedical Image SegmentationSemantic Segmentation

Segment Anything Model for Zero-shot Single Particle Tracking in Liquid Phase Transmission Electron Microscopy

2025-01-06 · Risha Goel, Zain Shabeeb, Isabel Panicker, Vida Jamali

Liquid phase transmission electron microscopy (LPTEM) offers an unparalleled combination of spatial and temporal resolution, making it a promising tool for single particle tracking at the nanoscale. However, the absence …

Video SegmentationVideo Semantic Segmentation

CUTS: A Deep Learning and Topological Framework for Multigranular Unsupervised Medical Image Segmentation

2022-09-23 · Chen Liu, Matthew Amodio, Liangbo L. Shen, Feng Gao 외

Segmenting medical images is critical to facilitating both patient diagnoses and quantitative research. A major limiting factor is the lack of labeled data, as obtaining expert annotations for each new set of imaging dat…

Contrastive LearningImage SegmentationMedical Image SegmentationSegmentation+4

Domain Adaptor Networks for Hyperspectral Image Recognition

2021-08-03 · Gustavo Perez, Subhransu Maji

We consider the problem of adapting a network trained on three-channel color images to a hyperspectral domain with a large number of channels. To this end, we propose domain adaptor networks that map the input to be comp…