paper-with-me

홈 › Papers

MAIS: Memory-Attention for Interactive Segmentation

2025-05-12 · Mauricio Orbes-Arteaga, Oeslle Lucena, Sabastien Ourselin, M. Jorge Cardoso

Interactive medical segmentation reduces annotation effort by refining predictions through user feedback. Vision Transformer (ViT)-based models, such as the Segment Anything Model (SAM), achieve state-of-the-art performance using user clicks and prior masks as prompts. However, existing methods treat interactions as independent events, leading to redundant corrections and limited refinement gains. We address this by introducing MAIS, a Memory-Attention mechanism for Interactive Segmentation that stores past user inputs and segmentation states, enabling temporal context integration. Our approach enhances ViT-based segmentation across diverse imaging modalities, achieving more efficient and accurate refinements.

📄 PDF Abstract BibTeX arXiv:2505.07511

Code (0)

등록된 구현이 없습니다.

Tasks

Interactive SegmentationSegmentation

Methods 이 논문이 사용한 방법론

Attention 설명 없음
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
Multi-Head Attention 설명 없음
Dense Connections Dense Connections, or Fully Connected Connections, are a type of layer in a deep neural network that use a linear operation where every input is connected to every output…
Vision Transformer The Vision Transformer, or ViT, is a model for image classification that employs a Transformer-like architecture over…
Dropout Dropout is a regularization technique for neural networks that drops a unit (along with connections) at training time with a specified probability $p$ (a common value is…
Layer Normalization Unlike batch normalization, Layer Normalization directly estimates the normalization statistics from the summed inputs…
Position-Wise Feed-Forward Layer 설명 없음

Similar Papers 제목 키워드 기반

TSDASeg: A Two-Stage Model with Direct Alignment for Interactive Point Cloud Segmentation

2025-06-26 · Chade Li, Pengju Zhang, Yihong Wu

The rapid advancement of 3D vision-language models (VLMs) has spurred significant interest in interactive point cloud processing tasks, particularly for real-world applications. However, existing methods often underperfo…

cross-modal alignmentInteractive SegmentationPoint Cloud SegmentationSegmentation+1

iSegFormer: Interactive Segmentation via Transformers with Application to 3D Knee MR Images

2021-12-21 · Qin Liu, Zhenlin Xu, Yining Jiao, Marc Niethammer

We propose iSegFormer, a memory-efficient transformer that combines a Swin transformer with a lightweight multilayer perceptron (MLP) decoder. With the efficient Swin transformer blocks for hierarchical self-attention an…

DecoderImage SegmentationInteractive SegmentationMedical Image Segmentation+1

MAISI: Medical AI for Synthetic Imaging

2024-09-13 · Pengfei Guo, Can Zhao, Dong Yang, Ziyue Xu 외

Medical imaging analysis faces challenges such as data scarcity, high annotation costs, and privacy concerns. This paper introduces the Medical AI for Synthetic Imaging (MAISI), an innovative approach using the diffusion…

Computed Tomography (CT)Organ Segmentation

mAIstro: an open-source multi-agentic system for automated end-to-end development of radiomics and deep learning models for medical imaging

2025-04-30 · Eleftherios Tzanis, Michail E. Klontzas

Agentic systems built on large language models (LLMs) offer promising capabilities for automating complex workflows in healthcare AI. We introduce mAIstro, an open-source, autonomous multi-agentic framework for end-to-en…

AI AgentClassificationLLM real-life tasksMedical Image Analysis+2

HRSAM: Efficient Interactive Segmentation in High-Resolution Images

2024-07-02 · You Huang, Wenbin Lai, Jiayi Ji, Liujuan Cao 외

The Segment Anything Model (SAM) has advanced interactive segmentation but is limited by the high computational cost on high-resolution images. This requires downsampling to meet GPU constraints, sacrificing the fine-gra…

Data AugmentationGPUInteractive SegmentationSegmentation+1