paper-with-me

홈 › Papers

Toward Zero-Shot Learning for Visual Dehazing of Urological Surgical Robots

2024-10-02 · Renkai Wu, Xianjin Wang, Pengchen Liang, Zhenyu Zhang, Qing Chang, Hao Tang

Robot-assisted surgery has profoundly influenced current forms of minimally invasive surgery. However, in transurethral suburethral urological surgical robots, they need to work in a liquid environment. This causes vaporization of the liquid when shearing and heating is performed, resulting in bubble atomization that affects the visual perception of the robot. This can lead to the need for uninterrupted pauses in the surgical procedure, which makes the surgery take longer. To address the atomization characteristics of liquids under urological surgical robotic vision, we propose an unsupervised zero-shot dehaze method (RSF-Dehaze) for urological surgical robotic vision. Specifically, the proposed Region Similarity Filling Module (RSFM) of RSF-Dehaze significantly improves the recovery of blurred region tissues. In addition, we organize and propose a dehaze dataset for robotic vision in urological surgery (USRobot-Dehaze dataset). In particular, this dataset contains the three most common urological surgical robot operation scenarios. To the best of our knowledge, we are the first to organize and propose a publicly available dehaze dataset for urological surgical robot vision. The proposed RSF-Dehaze proves the effectiveness of our method in three urological surgical robot operation scenarios with extensive comparative experiments with 20 most classical and advanced dehazing and image recovery algorithms. The proposed source code and dataset are available at https://github.com/wurenkai/RSF-Dehaze .

📄 PDF Abstract BibTeX arXiv:2410.01395

Code (1)

wurenkai/rsf-dehaze 공식 구현 pytorch

Tasks

Zero-Shot Learning

Similar Papers 제목 키워드 기반

Zero-shot Prompt-based Video Encoder for Surgical Gesture Recognition

2024-03-28 · Mingxing Rao, Yinhong Qin, Soheil Kolouri, Jie Ying Wu 외

Purpose: In order to produce a surgical gesture recognition system that can support a wide variety of procedures, either a very large annotated dataset must be acquired, or fitted models must generalize to new labels (so…

Gesture RecognitionSurgical Gesture Recognition

SurgCheck: Do Vision-Language Models Really Look at Images in Surgical VQA?

2026-05-03 · Jongmin Shin, Ka Young Kim, Eunki Cho, Seong Tae Kim 외 arxiv

Purpose: Vision-language models (VLMs) have shown promising performance in surgical visual question answering (VQA). However, existing surgical VQA datasets often contain linguistic shortcuts, where question phrasing imp…

Visual Question AnsweringVisual Reasoning

How Far Are Surgeons from Surgical World Models? A Pilot Study on Zero-shot Surgical Video Generation with Expert Assessment

2025-11-03 · Zhen Chen, Qing Xu, Jinlin Wu, Biao Yang 외 arxiv

Foundation models in video generation are demonstrating remarkable capabilities as potential world models for simulating the physical world. However, their application in high-stakes domains like surgery, which demand de…

Video Generation

SurgPose: Generalisable Surgical Instrument Pose Estimation using Zero-Shot Learning and Stereo Vision

2025-05-16 · Utsav Rai, Haozheng Xu, Stamatia Giannarou

Accurate pose estimation of surgical tools in Robot-assisted Minimally Invasive Surgery (RMIS) is essential for surgical navigation and robot control. While traditional marker-based methods offer accuracy, they face chal…

Depth EstimationInstance SegmentationPose EstimationSemantic Segmentation+1

HecVL: Hierarchical Video-Language Pretraining for Zero-shot Surgical Phase Recognition

2024-05-16 · Kun Yuan, Vinkle Srivastav, Nassir Navab, Nicolas Padoy

Natural language could play an important role in developing generalist surgical models by providing a broad source of supervision from raw texts. This flexible form of supervision can enable the model's transferability a…

Contrastive LearningSurgical phase recognition