paper-with-me

Papers

Language-guided Scale-aware MedSegmentor for Lesion Segmentation in Medical Imaging

2024-08-30 · Shuyi Ouyang, Jinyang Zhang, Xiangye Lin, Xilai Wang, Qingqing Chen, Yen-Wei Chen, Lanfen Lin

In clinical practice, segmenting specific lesions based on the needs of physicians can significantly enhance diagnostic accuracy and treatment efficiency. However, conventional lesion segmentation models lack the flexibility to distinguish lesions according to specific requirements. Given the practical advantages of using text as guidance, we propose a novel model, Language-guided Scale-aware MedSegmentor (LSMS), which segments target lesions in medical images based on given textual expressions. We define this as a new task termed Referring Lesion Segmentation (RLS). To address the lack of suitable benchmarks for RLS, we construct a vision-language medical dataset named Reference Hepatic Lesion Segmentation (RefHL-Seg). LSMS incorporates two key designs: (i) Scale-Aware Vision-Language attention module, which performs visual feature extraction and vision-language alignment in parallel. By leveraging diverse convolutional kernels, this module acquires rich visual representations and interacts closely with linguistic features, thereby enhancing the model's capacity for precise object localization. (ii) Full-Scale Decoder, which globally models multi-modal features across multiple scales and captures complementary information between them to accurately delineate lesion boundaries. Additionally, we design a specialized loss function comprising both segmentation loss and vision-language contrastive loss to better optimize cross-modal learning. We validate the performance of LSMS on RLS as well as on conventional lesion segmentation tasks across multiple datasets. Our LSMS consistently achieves superior performance with significantly lower computational cost. Code and datasets will be released.

📄 PDF Abstract BibTeX arXiv:2408.17347

Code (0)

등록된 구현이 없습니다.

Tasks

DiagnosticImage SegmentationLanguage ModelingLanguage ModellingLesion SegmentationMedical Image SegmentationObject LocalizationSegmentationSemantic Segmentation

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음

Similar Papers 제목 키워드 기반

GLeVE: Graph-Guided Lesion Grounding with Proposal Verification in 3D CT

2026-05-21 · Shuo Jiang, Yuhao Hong, Chunbo Jiang, Weihong Chen 외 arxiv

Grounding radiology report descriptions to 3D CT volumes is essential for verifiable clinical interpretation, yet remains challenging due to the semantic-spatial gap between free-text narratives and volumetric anatomy. E…

UHR-Net: An Uncertainty-Aware Hypergraph Refinement Network for Medical Image Segmentation

2026-04-30 · Shuokun Cheng, Jinghao Shi, Kun Sun arxiv

Accurate lesion segmentation is crucial for clinical diagnosis and treatment planning. However, lesions often resemble surrounding tissues and exhibit ill-defined boundaries, leading to unstable predictions in boundary/t…

Medical Image SegmentationLesion Segmentation

Region-Grounded Vision-Language Learning for Detection-Guided Mammographic Lesion Classification

2026-07-17 · Zhengbo Zhou, Jiren Li, Dooman Arefan, Margarita Zuley 외 arxiv

Vision-language models trained with contrastive objectives have shown promise in medical image analysis. However, conventional global image-text alignment is ill-suited for mammography, where diagnostically relevant lesi…

Transfer Learning

MedForge: Interpretable Medical Deepfake Detection via Forgery-aware Reasoning

2026-03-19 · Zhihui Chen, Kai He, Qingyuan Lei, Bin Pu 외 arxiv

Text-guided image editors can now manipulate authentic medical scans with high fidelity, enabling lesion implantation/removal that threatens clinical trust and safety. Existing defenses are inadequate for healthcare. Med…

DeepFake Detection

Heterogeneity-Adaptive Diffusion Schrodinger Bridge for PET-Guided Whole-Body MRI Translation

2026-07-08 · Chengbo Wang, Jiacheng Yu, Linjie Bian, Ming Qi 외 arxiv

While whole-body multimodal medical imaging scanners have been increasingly recognized for more effective medical applications, the excessive long acquisition time in PET-MR scanning is a major obstacle in more efficient…