paper-with-me

홈 › Papers

CLAPS: A CLIP-Unified Auto-Prompt Segmentation for Multi-Modal Retinal Imaging

2025-09-10 · Zhihao Zhao, Yinzheng Zhao, Junjie Yang, Xiangtong Yao, Quanmin Liang, Shahrooz Faghihroohi, Kai Huang, Nassir Navab, M. Ali Nasseri arxiv

Recent advancements in foundation models, such as the Segment Anything Model (SAM), have significantly impacted medical image segmentation, especially in retinal imaging, where precise segmentation is vital for diagnosis. Despite this progress, current methods face critical challenges: 1) modality ambiguity in textual disease descriptions, 2) a continued reliance on manual prompting for SAM-based workflows, and 3) a lack of a unified framework, with most methods being modality- and task-specific. To overcome these hurdles, we propose CLIP-unified Auto-Prompt Segmentation (\CLAPS), a novel method for unified segmentation across diverse tasks and modalities in retinal imaging. Our approach begins by pre-training a CLIP-based image encoder on a large, multi-modal retinal dataset to handle data scarcity and distribution imbalance. We then leverage GroundingDINO to automatically generate spatial bounding box prompts by detecting local lesions. To unify tasks and resolve ambiguity, we use text prompts enhanced with a unique "modality signature" for each imaging modality. Ultimately, these automated textual and spatial prompts guide SAM to execute precise segmentation, creating a fully automated and unified pipeline. Extensive experiments on 12 diverse datasets across 11 critical segmentation categories show that CLAPS achieves performance on par with specialized expert models while surpassing existing benchmarks across most metrics, demonstrating its broad generalizability as a foundation model.

📄 PDF Abstract BibTeX arXiv:2509.08618

Code (0)

등록된 구현이 없습니다.

Tasks

Medical Image Segmentation

Similar Papers 제목 키워드 기반

Test-Time Adaptation with SaLIP: A Cascade of SAM and CLIP for Zero shot Medical Image Segmentation

2024-04-09 · Sidra Aleem, Fangyijie Wang, Mayug Maniparambil, Eric Arazo 외

The Segment Anything Model (SAM) and CLIP are remarkable vision foundation models (VFMs). SAM, a prompt driven segmentation model, excels in segmentation tasks across diverse domains, while CLIP is renowned for its zero …

Image SegmentationMedical Image SegmentationOrgan SegmentationPrompt Engineering+5

CLAPSep: Leveraging Contrastive Pre-trained Model for Multi-Modal Query-Conditioned Target Sound Extraction

2024-02-27 · Hao Ma, Zhiyuan Peng, Xu Li, Mingjie Shao 외

Universal sound separation (USS) aims to extract arbitrary types of sounds from real-world recordings. This can be achieved by language-queried target sound extraction (TSE), which typically consists of two components: a…

Target Sound Extraction

PointCLIP V2: Prompting CLIP and GPT for Powerful 3D Open-world Learning

2022-11-21 · ICCV 2023 1 · Xiangyang Zhu, Renrui Zhang, Bowei He, Ziyu Guo 외

Large-scale pre-trained models have shown promising open-world performance for both vision and language tasks. However, their transferred capacity on 3D point clouds is still limited and only constrained to the classific…

3D Classification3D Object Detection3D Open-Vocabulary Instance Segmentation3D Part Segmentation+10

VCP-CLIP: A visual context prompting model for zero-shot anomaly segmentation

2024-07-17 · Zhen Qu, Xian Tao, Mukesh Prasad, Fei Shen 외

Recently, large-scale vision-language models such as CLIP have demonstrated immense potential in zero-shot anomaly segmentation (ZSAS) task, utilizing a unified model to directly detect anomalies on any unseen product wi…

Anomaly DetectionAnomaly Segmentationzero-shot anomaly detectionZero Shot Segmentation

ClipSAM: CLIP and SAM Collaboration for Zero-Shot Anomaly Segmentation

2024-01-23 · Shengze Li, JianJian Cao, Peng Ye, Yuhan Ding 외

Recently, foundational models such as CLIP and SAM have shown promising performance for the task of Zero-Shot Anomaly Segmentation (ZSAS). However, either CLIP-based or SAM-based ZSAS methods still suffer from non-neglig…

Anomaly LocalizationAnomaly SegmentationSegmentation