paper-with-me

홈 › Papers

Fine-Grained Action Segmentation for Renorrhaphy in Robot-Assisted Partial Nephrectomy

2026-04-10 · Jiaheng Dai, Huanrong Liu, Tailai Zhou, Tongyu Jia, Qin Liu, Yutong Ban, Zeju Li, Yu Gao, Xin Ma, Qingbiao Li arxiv

Fine-grained action segmentation during renorrhaphy in robot-assisted partial nephrectomy requires frame-level recognition of visually similar suturing gestures with variable duration and substantial class imbalance. The SIA-RAPN benchmark defines this problem on 50 clinical videos acquired with the da Vinci Xi system and annotated with 12 frame-level labels. The benchmark compares four temporal models built on I3D features: MS-TCN++, AsFormer, TUT, and DiffAct. Evaluation uses balanced accuracy, edit score, segmental F1 at overlap thresholds of 10, 25, and 50, frame-wise accuracy, and frame-wise mean average precision. In addition to the primary evaluation across five released split configurations on SIA-RAPN, the benchmark reports cross-domain results on a separate single-port RAPN dataset. Across the strongest reported values over those five runs on the primary dataset, DiffAct achieves the highest F1, frame-wise accuracy, edit score, and frame mAP, while MS-TCN++ attains the highest balanced accuracy.

📄 PDF Abstract BibTeX arXiv:2604.09051

Code (0)

등록된 구현이 없습니다.

Tasks

Action Segmentation

Similar Papers 제목 키워드 기반

Hand Guided High Resolution Feature Enhancement for Fine-Grained Atomic Action Segmentation within Complex Human Assemblies

2022-11-24 · Matthew Kent Myers, Nick Wright, Stephen McGough, Nicholas Martin

Due to the rapid temporal and fine-grained nature of complex human assembly atomic actions, traditional action segmentation approaches requiring the spatial (and often temporal) down sampling of video frames often loose …

Action ClassificationAction RecognitionAction SegmentationClassification+2

Segmental Spatiotemporal CNNs for Fine-grained Action Segmentation

2016-02-09 · Colin Lea, Austin Reiter, Rene Vidal, Gregory D. Hager

Joint segmentation and classification of fine-grained actions is important for applications of human-robot interaction, video surveillance, and human skill evaluation. However, despite substantial recent progress in larg…

Action ClassificationAction RecognitionAction SegmentationFine-grained Action Recognition+3

Technical Report for ICRA 2026 GOOSE 2D Fine-Grained Semantic Segmentation Challenge: Exploring Query-Based Segmentation and Increased Spatial Context for Outdoor Scene Understanding

2026-06-19 · David Pascual-Hernández, Roberto Calvo-Palomino, Inmaculada Mora-Jiménez, Jose María Cañas-Plaza arxiv

In this report, we present our submission to the GOOSE 2D Fine-Grained Semantic Segmentation Challenge, organized as part of the Workshop on Field Robotics at ICRA 2026. The challenge combines data from the GOOSE and GOO…

Semantic SegmentationScene Understanding

Temporal Convolutional Networks for Action Segmentation and Detection

2016-11-16 · CVPR 2017 7 · Colin Lea, Michael D. Flynn, Rene Vidal, Austin Reiter 외

The ability to identify and temporally segment fine-grained human actions throughout a video is crucial for robotics, surveillance, education, and beyond. Typical approaches decouple this problem by first extracting loca…

Action SegmentationDecoderSkeleton Based Action Recognition

Kaiwu: A Multimodal Manipulation Dataset and Framework for Robot Learning and Human-Robot Interaction

2025-03-07 · Shuo Jiang, Haonan Li, Ruochen Ren, Yanmin Zhou 외

Cutting-edge robot learning techniques including foundation models and imitation learning from humans all pose huge demands on large-scale and high-quality datasets which constitute one of the bottleneck in the general i…

Imitation LearningSemantic Segmentation