paper-with-me

Papers

Self-Supervised Contrastive Embedding Adaptation for Endoscopic Image Matching

2025-12-11 · Alberto Rota, Elena De Momi arxiv

Accurate spatial understanding is essential for image-guided surgery, augmented reality integration and context awareness. In minimally invasive procedures, where visual input is the sole intraoperative modality, establishing precise pixel-level correspondences between endoscopic frames is critical for 3D reconstruction, camera tracking, and scene interpretation. However, the surgical domain presents distinct challenges: weak perspective cues, non-Lambertian tissue reflections, and complex, deformable anatomy degrade the performance of conventional computer vision techniques. While Deep Learning models have shown strong performance in natural scenes, their features are not inherently suited for fine-grained matching in surgical images and require targeted adaptation to meet the demands of this domain. This research presents a novel Deep Learning pipeline for establishing feature correspondences in endoscopic image pairs, alongside a self-supervised optimization framework for model training. The proposed methodology leverages a novel-view synthesis pipeline to generate ground-truth inlier correspondences, subsequently utilized for mining triplets within a contrastive learning paradigm. Through this self-supervised approach, we augment the DINOv2 backbone with an additional Transformer layer, specifically optimized to produce embeddings that facilitate direct matching through cosine similarity thresholding. Experimental evaluation demonstrates that our pipeline surpasses state-of-the-art methodologies on the SCARED datasets improved matching precision and lower epipolar error compared to the related work. The proposed framework constitutes a valuable contribution toward enabling more accurate high-level computer vision applications in surgical endoscopy.

📄 PDF Abstract BibTeX arXiv:2512.10379

Code (0)

등록된 구현이 없습니다.

Tasks

Contrastive Learning3D ReconstructionImage Matching

Similar Papers 제목 키워드 기반

EndoDAC: Efficient Adapting Foundation Model for Self-Supervised Depth Estimation from Any Endoscopic Camera

2024-05-14 · Beilei Cui, Mobarakol Islam, Long Bai, An Wang 외

Depth estimation plays a crucial role in various tasks within endoscopic surgery, including navigation, surface reconstruction, and augmented reality visualization. Despite the significant achievements of foundation mode…

Depth EstimationSurface Reconstruction

Acoustic word embeddings for zero-resource languages using self-supervised contrastive learning and multilingual adaptation

2021-03-19 · Christiaan Jacobs, Yevgen Matusevych, Herman Kamper

Acoustic word embeddings (AWEs) are fixed-dimensional representations of variable-length speech segments. For zero-resource languages where labelled data is not available, one AWE approach is to use unsupervised autoenco…

Contrastive LearningWord Embeddings

QA Domain Adaptation using Hidden Space Augmentation and Self-Supervised Contrastive Adaptation

2022-10-19 · Zhenrui Yue, Huimin Zeng, Bernhard Kratzwald, Stefan Feuerriegel 외

Question answering (QA) has recently shown impressive results for answering questions from customized domains. Yet, a common challenge is to adapt QA models to an unseen target domain. In this paper, we propose a novel s…

Contrastive LearningData AugmentationDomain AdaptationQuestion Answering

MetaFE-DE: Learning Meta Feature Embedding for Depth Estimation from Monocular Endoscopic Images

2025-02-05 · Dawei Lu, Deqiang Xiao, Danni Ai, Jingfan Fan 외

Depth estimation from monocular endoscopic images presents significant challenges due to the complexity of endoscopic surgery, such as irregular shapes of human soft tissues, as well as variations in lighting conditions.…

Depth EstimationMonocular Depth EstimationSelf-Supervised Learning

Endo-CLIP: Progressive Self-Supervised Pre-training on Raw Colonoscopy Records

2025-05-14 · Yili He, Yan Zhu, Peiyao Fu, Ruijie Yang 외

Pre-training on image-text colonoscopy records offers substantial potential for improving endoscopic image analysis, but faces challenges including non-informative background images, complex medical terminology, and ambi…

Contrastive Learning