paper-with-me

홈 › Papers

Affordance-Based Disambiguation of Surgical Instructions for Collaborative Robot-Assisted Surgery

2025-09-18 · Ana Davila, Jacinto Colan, Yasuhisa Hasegawa arxiv

Effective human-robot collaboration in surgery is affected by the inherent ambiguity of verbal communication. This paper presents a framework for a robotic surgical assistant that interprets and disambiguates verbal instructions from a surgeon by grounding them in the visual context of the operating field. The system employs a two-level affordance-based reasoning process that first analyzes the surgical scene using a multimodal vision-language model and then reasons about the instruction using a knowledge base of tool capabilities. To ensure patient safety, a dual-set conformal prediction method is used to provide a statistically rigorous confidence measure for robot decisions, allowing it to identify and flag ambiguous commands. We evaluated our framework on a curated dataset of ambiguous surgical requests from cholecystectomy videos, demonstrating a general disambiguation rate of 60% and presenting a method for safer human-robot interaction in the operating room.

📄 PDF Abstract BibTeX arXiv:2509.14967

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

LLM-based ambiguity detection in natural language instructions for collaborative surgical robots

2025-07-15 · Ana Davila, Jacinto Colan, Yasuhisa Hasegawa arxiv

Ambiguity in natural language instructions poses significant risks in safety-critical human-robot interaction, particularly in domains such as surgery. To address this, we propose a framework that uses Large Language Mod…

arg-VU: Affordance Reasoning with Physics-Aware 3D Geometry for Visual Understanding in Robotic Surgery

2026-03-26 · Nan Xiao, Yunxin Fan, Farong Wang, Fei Liu arxiv

Affordance reasoning provides a principled link between perception and action, yet remains underexplored in surgical robotics, where tissues are highly deformable, compliant, and dynamically coupled with tool motion. We …

RAGNet: Large-scale Reasoning-based Affordance Segmentation Benchmark towards General Grasping

2025-07-31 · Dongming Wu, Yanping Fu, Saike Huang, Yingfei Liu 외 arxiv

General robotic grasping systems require accurate object affordance perception in diverse open-world scenarios following human instructions. However, current studies suffer from the problem of lacking reasoning-based lar…

Robot ManipulationRobotic Grasping

GREAT: Geometry-Intention Collaborative Inference for Open-Vocabulary 3D Object Affordance Grounding

2024-11-29 · CVPR 2025 1 · Yawen Shao, Wei Zhai, Yuhang Yang, Hongchen Luo 외

Open-Vocabulary 3D object affordance grounding aims to anticipate ``action possibilities'' regions on 3D objects with arbitrary instructions, which is crucial for robots to generically perceive real scenarios and respond…

Collaborative InferenceObject

SurgAM: Surgical Affordance Map Prediction with Multimodal Feature Fusion for Robot Autonomy

2026-07-05 · Lei Song, Yonghao Long, Mengya Xu, Jiayi Geng 외 arxiv

Surgical automation is being increasingly studied, yet bridging visual scene understanding with autonomous action planning remains a fundamental challenge. While much research effort has been made on scene perception (e.…

Scene UnderstandingScene Segmentation