paper-with-me

홈 › Papers

CAT: Coordinating Anatomical-Textual Prompts for Multi-Organ and Tumor Segmentation

2024-06-11 · Zhongzhen Huang, Yankai Jiang, Rongzhao Zhang, Shaoting Zhang, Xiaofan Zhang

Existing promptable segmentation methods in the medical imaging field primarily consider either textual or visual prompts to segment relevant objects, yet they often fall short when addressing anomalies in medical images, like tumors, which may vary greatly in shape, size, and appearance. Recognizing the complexity of medical scenarios and the limitations of textual or visual prompts, we propose a novel dual-prompt schema that leverages the complementary strengths of visual and textual prompts for segmenting various organs and tumors. Specifically, we introduce CAT, an innovative model that Coordinates Anatomical prompts derived from 3D cropped images with Textual prompts enriched by medical domain knowledge. The model architecture adopts a general query-based design, where prompt queries facilitate segmentation queries for mask prediction. To synergize two types of prompts within a unified framework, we implement a ShareRefiner, which refines both segmentation and prompt queries while disentangling the two types of prompts. Trained on a consortium of 10 public CT datasets, CAT demonstrates superior performance in multiple segmentation tasks. Further validation on a specialized in-house dataset reveals the remarkable capacity of segmenting tumors across multiple cancer stages. This approach confirms that coordinating multimodal prompts is a promising avenue for addressing complex scenarios in the medical domain.

📄 PDF Abstract BibTeX arXiv:2406.07085

Code (1)

zongzi3zz/cat 공식 구현 pytorch

Tasks

SegmentationTumor Segmentation

Similar Papers 제목 키워드 기반

Organ-aware Multi-scale Medical Image Segmentation Using Text Prompt Engineering

2025-03-18 · Wenjie Zhang, Ziyang Zhang, Mengnan He, Jiancheng Ye

Accurate segmentation is essential for effective treatment planning and disease monitoring. Existing medical image segmentation methods predominantly rely on uni-modal visual inputs, such as images or videos, requiring l…

BenchmarkingDescriptiveImage SegmentationMedical Image Segmentation+4

Text-promptable Propagation for Referring Medical Image Sequence Segmentation

2025-02-16 · Runtian Yuan, Jilan Xu, Mohan Chen, Qingqiu Li 외

Medical image sequences, generated by both 2D video-based examinations and 3D imaging techniques, consist of sequential frames or slices that capture the same anatomical entities (e.g., organs or lesions) from multiple p…

Interactive SegmentationSegmentation

GuideGen: A Text-Guided Framework for Full-torso Anatomy and CT Volume Generation

2024-03-12 · Linrui Dai, Rongzhao Zhang, Yongrui Yu, Xiaofan Zhang

The recently emerging conditional diffusion models seem promising for mitigating the labor and expenses in building large 3D medical imaging datasets. However, previous studies on 3D CT generation have yet to fully capit…

AnatomyTumor Segmentation

Leveraging Textual Anatomical Knowledge for Class-Imbalanced Semi-Supervised Multi-Organ Segmentation

2025-01-23 · Yuliang Gu, Weilun Tsao, Bo Du, Thierry Géraud 외

Annotating 3D medical images demands substantial time and expertise, driving the adoption of semi-supervised learning (SSL) for segmentation tasks. However, the complex anatomical structures of organs often lead to signi…

Contrastive LearningOrgan SegmentationSegmentation

MOSAIC: A Multi-View 2.5D Organ Slice Selector with Cross-Attentional Reasoning for Anatomically-Aware CT Localization in Medical Organ Segmentation

2025-05-15 · Hania Ghouse, Muzammil Behzad

Efficient and accurate multi-organ segmentation from abdominal CT volumes is a fundamental challenge in medical image analysis. Existing 3D segmentation approaches are computationally and memory intensive, often processi…

Medical Image AnalysisOrgan SegmentationSegmentation