paper-with-me

홈 › Papers

SAM2CLIP2SAM: Vision Language Model for Segmentation of 3D CT Scans for Covid-19 Detection

2024-07-22 · Dimitrios Kollias, Anastasios Arsenos, James Wingate, Stefanos Kollias

This paper presents a new approach for effective segmentation of images that can be integrated into any model and methodology; the paradigm that we choose is classification of medical images (3-D chest CT scans) for Covid-19 detection. Our approach includes a combination of vision-language models that segment the CT scans, which are then fed to a deep neural architecture, named RACNet, for Covid-19 detection. In particular, a novel framework, named SAM2CLIP2SAM, is introduced for segmentation that leverages the strengths of both Segment Anything Model (SAM) and Contrastive Language-Image Pre-Training (CLIP) to accurately segment the right and left lungs in CT scans, subsequently feeding these segmented outputs into RACNet for classification of COVID-19 and non-COVID-19 cases. At first, SAM produces multiple part-based segmentation masks for each slice in the CT scan; then CLIP selects only the masks that are associated with the regions of interest (ROIs), i.e., the right and left lungs; finally SAM is given these ROIs as prompts and generates the final segmentation mask for the lungs. Experiments are presented across two Covid-19 annotated databases which illustrate the improved performance obtained when our method has been used for segmentation of the CT scans.

📄 PDF Abstract BibTeX arXiv:2407.15728

Code (0)

등록된 구현이 없습니다.

Tasks

Language ModelingLanguage ModellingSegmentation

Methods 이 논문이 사용한 방법론

CLIP Contrastive Language-Image Pre-training (CLIP), consisting of a simplified version of ConVIRT trained from scratch, is an efficient method of image representation learning…
SAM 설명 없음

Similar Papers 제목 키워드 기반

MedFocusCLIP : Improving few shot classification in medical datasets using pixel wise attention

2025-01-07 · Aadya Arora, Vinay Namboodiri

With the popularity of foundational models, parameter efficient fine tuning has become the defacto approach to leverage pretrained models to perform downstream tasks. Taking inspiration from recent advances in large lang…

ClassificationFine-Grained Image Classificationimage-classificationImage Classification+4

Interactive Segmentation for COVID-19 Infection Quantification on Longitudinal CT scans

2021-10-03 · Michelle Xiao-Lin Foo, Seong Tae Kim, Magdalini Paschali, Leili Goli 외

Consistent segmentation of COVID-19 patient's CT scans across multiple time points is essential to assess disease progression and response to therapy accurately. Existing automatic and interactive segmentation models for…

Interactive SegmentationSegmentation

Detection and Segmentation of Lesion Areas in Chest CT Scans For The Prediction of COVID-19

2020-10-26 · Aram Ter-Sarkisov

In this paper we compare the models for the detection and segmentation of Ground Glass Opacity and Consolidation in chest CT scans. These lesion areas are often associated both with common pneumonia and COVID-19. We trai…

COVID-19 DiagnosisCOVID-19 Image SegmentationInstance SegmentationSegmentation+2

COVID-FACT: A Fully-Automated Capsule Network-based Framework for Identification of COVID-19 Cases from Chest CT scans

2020-10-30 · Shahin Heidarian, Parnian Afshar, Nastaran Enshaei, Farnoosh Naderkhani 외

The newly discovered Corona virus Disease 2019 (COVID-19) has been globally spreading and causing hundreds of thousands of deaths around the world as of its first emergence in late 2019. Computed tomography (CT) scans ha…

Computed Tomography (CT)Data AugmentationDiagnosticSensitivity+1

Single-Shot Lightweight Model For The Detection of Lesions And The Prediction of COVID-19 From Chest CT Scans

2020-12-02 · Aram Ter-Sarkisov

We introduce a lightweight model based on Mask R-CNN with ResNet18 and ResNet34 backbone models that segments lesions and predicts COVID-19 from chest CT scans in a single shot. The model requires a small dataset to tr…

ClassificationCOVID-19 DiagnosisGeneral ClassificationInstance Segmentation+3