paper-with-me

홈 › Papers

Cataract-LMM Large-Scale Multi-Source Multi-Task Benchmark for Deep Learning in Surgical Video Analysis

2025-10-18 · Mohammad Javad Ahmadi, Iman Gandomi, Parisa Abdi, Seyed-Farzad Mohammadi, Amirhossein Taslimi, Mehdi Khodaparast, Hassan Hashemi, Mahdi Tavakoli, Hamid D. Taghirad arxiv

Computer-assisted surgery research requires large, deeply annotated video datasets that capture clinical and technical variability. Existing cataract surgery resources lack the diversity and annotation depth required to train generalizable deep-learning models. To address this gap, we present a dataset of 3,000 phacoemulsification cataract surgery videos acquired at two surgical centers from surgeons with varying expertise. The dataset provides four annotation layers: temporal surgical phases, instance segmentation of instruments and anatomical structures, instrument-tissue interaction tracking, and quantitative skill scores based on competency rubrics adapted from ICO-OSCAR and GRASIS. We demonstrate the technical utility of the dataset through benchmarking deep learning models across four tasks: workflow recognition, scene segmentation, instrument-tissue interaction tracking, and automated skill assessment. Furthermore, we establish a domain-adaptation baseline for phase recognition and instance segmentation by training on one surgical center and evaluating on a held-out center. Ultimately, these multi-source acquisitions, multi-layer annotations, and paired skill-kinematic labels facilitate the development of generalizable multi-task models for surgical workflow analysis, scene understanding, and competency-based training research.

📄 PDF Abstract BibTeX arXiv:2510.16371

Code (0)

등록된 구현이 없습니다.

Tasks

Instance SegmentationScene UnderstandingScene Segmentation

Similar Papers 제목 키워드 기반

Video-Based Detection of squint and cataract for accessibility-aware adaptive web interface rendering

2026-07-08 · Amar Ranjan Dash, Manas Ranjan Patra arxiv

Squint and cataract are major ocular disorders that majorly affect visual perception and interaction capability. This paper proposes a real-time video-based automated detection system for squint and cataract detection ba…

Facial Landmark Detection

CataractBot: An LLM-Powered Expert-in-the-Loop Chatbot for Cataract Patients

2024-02-07 · Pragnya Ramjee, Bhuvan Sachdeva, Satvik Golechha, Shreyas Kulkarni 외

The healthcare landscape is evolving, with patients seeking reliable information about their health conditions and available treatment options. Despite the abundance of information sources, the digital age overwhelms ind…

Chatbot

MTCD: Cataract Detection via Near Infrared Eye Images

2021-10-06 · Pavani Tripathi, Yasmeena Akhter, Mahapara Khurshid, Aditya Lakra 외

Globally, cataract is a common eye disease and one of the leading causes of blindness and vision impairment. The traditional process of detecting cataracts involves eye examination using a slit-lamp microscope or ophthal…

ClassificationIris Recognition

Cataract-1K: Cataract Surgery Dataset for Scene Segmentation, Phase Recognition, and Irregularity Detection

2023-12-11 · Negin Ghamsarian, Yosuf El-Shabrawi, Sahar Nasirihaghighi, Doris Putzgruber-Adamitsch 외

In recent years, the landscape of computer-assisted interventions and post-operative surgical video analysis has been dramatically reshaped by deep-learning techniques, resulting in significant advancements in surgeons' …

BenchmarkingDomain AdaptationManagementScene Segmentation+2

Predicting Postoperative Intraocular Lens Dislocation in Cataract Surgery via Deep Learning

2023-12-06 · Negin Ghamsarian, Doris Putzgruber-Adamitsch, Stephanie Sarny, Raphael Sznitman 외

A critical yet unpredictable complication following cataract surgery is intraocular lens dislocation. Postoperative stability is imperative, as even a tiny decentration of multifocal lenses or inadequate alignment of the…