paper-with-me

Papers

Rethinking Text-Promptable Surgical Instrument Segmentation with Robust Framework

2024-11-19 · Tae-Min Choi, Juyoun Park

Surgical instrument segmentation is an essential component of computer-assisted and robotic surgery systems. Vision-based segmentation models typically produce outputs limited to a predefined set of instrument categories, which restricts their applicability in interactive systems and robotic task automation. Promptable segmentation methods allow selective predictions based on textual prompts. However, they often rely on the assumption that the instruments present in the scene are already known, and prompts are generated accordingly, limiting their ability to generalize to unseen or dynamically emerging instruments. In practical surgical environments, where instrument existence information is not provided, this assumption does not hold consistently, resulting in false-positive segmentation. To address these limitations, we formulate a new task called Robust text-promptable Surgical Instrument Segmentation (R-SIS). Under this setting, prompts are issued for all candidate categories without access to instrument presence information. R-SIS requires distinguishing which prompts refer to visible instruments and generating masks only when such instruments are explicitly present in the scene. This setting reflects practical conditions where uncertainty in instrument presence is inherent. We evaluate existing segmentation methods under the R-SIS protocol using surgical video datasets and observe substantial false-positive predictions in the absence of ground-truth instruments. These findings demonstrate a mismatch between current evaluation protocols and real-world use cases, and support the need for benchmarks that explicitly account for prompt uncertainty and instrument absence.

📄 PDF Abstract BibTeX arXiv:2411.12199

Code (0)

등록된 구현이 없습니다.

Tasks

Segmentation

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

Text Promptable Surgical Instrument Segmentation with Vision-Language Models

2023-06-15 · NeurIPS 2023 11

In this paper, we propose a novel text promptable surgical instrument segmentation approach to overcome challenges associated with diversity and differentiation of surgical instruments in minimally invasive surgeries. We…

DecoderDiversitySegmentation

SurgicalPart-SAM: Part-to-Whole Collaborative Prompting for Surgical Instrument Segmentation

2023-12-22 · Wenxi Yue, Jing Zhang, Kun Hu, Qiuxia Wu 외

The Segment Anything Model (SAM) exhibits promise in generic object segmentation and offers potential for various applications. Existing methods have applied SAM to surgical instrument segmentation (SIS) by tuning SAM-ba…

SegmentationSemantic Segmentation

SurgicalSAM: Efficient Class Promptable Surgical Instrument Segmentation

2023-08-17 · Wenxi Yue, Jing Zhang, Kun Hu, Yong Xia 외

The Segment Anything Model (SAM) is a powerful foundation model that has revolutionised image segmentation. To apply SAM to surgical instrument segmentation, a common approach is to locate precise points or boxes of inst…

Image SegmentationSegmentationSemantic Segmentation

Rethinking Surgical Instrument Segmentation: A Background Image Can Be All You Need

2022-06-23 · An Wang, Mobarakol Islam, Mengya Xu, Hongliang Ren

Data diversity and volume are crucial to the success of training deep learning models, while in the medical imaging field, the difficulty and cost of data collection and annotation are especially huge. Specifically in ro…

AllDeep LearningDomain AdaptationIncremental Learning+2

SurgTPGS: Semantic 3D Surgical Scene Understanding with Text Promptable Gaussian Splatting

2025-06-29 · Yiming Huang, Long Bai, Beilei Cui, Kun Yuan 외

In contemporary surgical research and practice, accurately comprehending 3D surgical scenes with text-promptable capabilities is particularly crucial for surgical planning and real-time intra-operative guidance, where pr…

3D ReconstructionScene Understanding