paper-with-me

Papers

m2caiSeg: Semantic Segmentation of Laparoscopic Images using Convolutional Neural Networks

2020-08-23 · Salman Maqbool, Aqsa Riaz, Hasan Sajid, Osman Hasan

Autonomous surgical procedures, in particular minimal invasive surgeries, are the next frontier for Artificial Intelligence research. However, the existing challenges include precise identification of the human anatomy and the surgical settings, and modeling the environment for training of an autonomous agent. To address the identification of human anatomy and the surgical settings, we propose a deep learning based semantic segmentation algorithm to identify and label the tissues and organs in the endoscopic video feed of the human torso region. We present an annotated dataset, m2caiSeg, created from endoscopic video feeds of real-world surgical procedures. Overall, the data consists of 307 images, each of which is annotated for the organs and different surgical instruments present in the scene. We propose and train a deep convolutional neural network for the semantic segmentation task. To cater for the low quantity of annotated data, we use unsupervised pre-training and data augmentation. The trained model is evaluated on an independent test set of the proposed dataset. We obtained a F1 score of 0.33 while using all the labeled categories for the semantic segmentation task. Secondly, we labeled all instruments into an 'Instruments' superclass to evaluate the model's performance on discerning the various organs and obtained a F1 score of 0.57. We propose a new dataset and a deep learning method for pixel level identification of various organs and instruments in a endoscopic surgical scene. Surgical scene understanding is one of the first steps towards automating surgical procedures.

📄 PDF Abstract BibTeX arXiv:2008.10134

Code (1)

salmanmaq/segmentationNetworks 공식 구현 pytorch

Tasks

AnatomyData AugmentationScene UnderstandingSegmentationSemantic SegmentationUnsupervised Pre-training

Methods 이 논문이 사용한 방법론

Kaiming Initialization 설명 없음
Max Pooling Max Pooling is a pooling operation that calculates the maximum value for patches of a feature map, and uses it to create a downsampled (pooled) feature map. It is usually…
Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
ReLU How Do I Communicate to Expedia? How Do I Communicate to Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Live Support & Special Travel…
Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…
Batch Normalization 설명 없음
SegNet SegNet is a semantic segmentation model. This core trainable segmentation architecture consists of an encoder network, a corresponding decoder network followed by a pixel-wise…

Similar Papers 제목 키워드 기반

Synthetic and Real Inputs for Tool Segmentation in Robotic Surgery

2020-07-17 · Emanuele Colleoni, Philip Edwards, Danail Stoyanov

Semantic tool segmentation in surgical videos is important for surgical scene understanding and computer-assisted interventions as well as for the development of robotic automation. The problem is challenging because dif…

Deep LearningScene UnderstandingSegmentation

CholecSeg8k: A Semantic Segmentation Dataset for Laparoscopic Cholecystectomy Based on Cholec80

2020-12-23 · W. -Y. Hong, C. -L. Kao, Y. -H. Kuo, J. -R. Wang 외

Computer-assisted surgery has been developed to enhance surgery correctness and safety. However, researchers and engineers suffer from limited annotated data to develop and train better algorithms. Consequently, the deve…

Semantic SegmentationSimultaneous Localization and Mapping

Unsupervised temporal context learning using convolutional neural networks for laparoscopic workflow analysis

2017-02-13 · Sebastian Bodenstedt, Martin Wagner, Darko Katić, Patrick Mietkowski 외

Computer-assisted surgery (CAS) aims to provide the surgeon with the right type of assistance at the right moment. Such assistance systems are especially relevant in laparoscopic surgery, where CAS can alleviate some of …

LapFM: A Laparoscopic Segmentation Foundation Model via Hierarchical Concept Evolving Pre-training

2025-12-09 · Qing Xu, Kun Yuan, Yuxiang Luo, Yuhao Zhai 외 arxiv

Surgical segmentation is pivotal for scene understanding yet remains hindered by annotation scarcity and semantic inconsistency across diverse procedures. Existing approaches typically fine-tune natural foundation models…

Scene Understanding

Generating large labeled data sets for laparoscopic image processing tasks using unpaired image-to-image translation

2019-07-05 · Micha Pfeiffer, Isabel Funke, Maria R. Robu, Sebastian Bodenstedt 외

In the medical domain, the lack of large training data sets and benchmarks is often a limiting factor for training deep neural networks. In contrast to expensive manual labeling, computer simulations can generate large a…

Image-to-Image TranslationLiver SegmentationTranslationvalid