paper-with-me

Papers

Endo-TTAP: Robust Endoscopic Tissue Tracking via Multi-Facet Guided Attention and Hybrid Flow-point Supervision

2025-03-28 · Rulin Zhou, Wenlong He, An Wang, Qiqi Yao, Haijun Hu, Jiankun Wang, Xi Zhang an Hongliang Ren

Accurate tissue point tracking in endoscopic videos is critical for robotic-assisted surgical navigation and scene understanding, but remains challenging due to complex deformations, instrument occlusion, and the scarcity of dense trajectory annotations. Existing methods struggle with long-term tracking under these conditions due to limited feature utilization and annotation dependence. We present Endo-TTAP, a novel framework addressing these challenges through: (1) A Multi-Facet Guided Attention (MFGA) module that synergizes multi-scale flow dynamics, DINOv2 semantic embeddings, and explicit motion patterns to jointly predict point positions with uncertainty and occlusion awareness; (2) A two-stage curriculum learning strategy employing an Auxiliary Curriculum Adapter (ACA) for progressive initialization and hybrid supervision. Stage I utilizes synthetic data with optical flow ground truth for uncertainty-occlusion regularization, while Stage II combines unsupervised flow consistency and semi-supervised learning with refined pseudo-labels from off-the-shelf trackers. Extensive validation on two MICCAI Challenge datasets and our collected dataset demonstrates that Endo-TTAP achieves state-of-the-art performance in tissue point tracking, particularly in scenarios characterized by complex endoscopic conditions. The source code and dataset will be available at https://anonymous.4open.science/r/Endo-TTAP-36E5.

📄 PDF Abstract BibTeX arXiv:2503.22394

Code (0)

등록된 구현이 없습니다.

Tasks

Optical Flow EstimationPoint TrackingScene Understanding

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음
Adapter 설명 없음

Similar Papers 제목 키워드 기반

EndoGSLAM: Real-Time Dense Reconstruction and Tracking in Endoscopic Surgeries using Gaussian Splatting

2024-03-22 · Kailing Wang, Chen Yang, Yuehao Wang, Sikuang Li 외

Precise camera tracking, high-fidelity 3D tissue reconstruction, and real-time online visualization are critical for intrabody medical imaging devices such as endoscopes and capsule robots. However, existing SLAM (Simult…

Simultaneous Localization and Mapping

Semantic-SuPer: A Semantic-aware Surgical Perception Framework for Endoscopic Tissue Identification, Reconstruction, and Tracking

2022-10-29 · Shan Lin, Albert J. Miao, Jingpei Lu, Shunkai Yu 외

Accurate and robust tracking and reconstruction of the surgical scene is a critical enabling technology toward autonomous robotic surgery. Existing algorithms for 3D perception in surgery mainly rely on geometric informa…

3D ReconstructionImage SegmentationSemantic Segmentation

Tracking Any Point Methods for Markerless 3D Tissue Tracking in Endoscopic Stereo Images

2025-08-11 · Konrad Reuter, Suresh Guttikonda, Sarah Latus, Lennart Maack 외 arxiv

Minimally invasive surgery presents challenges such as dynamic tissue motion and a limited field of view. Accurate tissue tracking has the potential to support surgical guidance, improve safety by helping avoid damage to…

FLex: Joint Pose and Dynamic Radiance Fields Optimization for Stereo Endoscopic Videos

2024-03-18 · Florian Philipp Stilz, Mert Asim Karaoglu, Felix Tristram, Nassir Navab 외

Reconstruction of endoscopic scenes is an important asset for various medical applications, from post-surgery analysis to educational training. Neural rendering has recently shown promising results in endoscopic reconstr…

Neural RenderingNovel View Synthesis

EndoGaussian: Real-time Gaussian Splatting for Dynamic Endoscopic Scene Reconstruction

2024-01-23 · Yifan Liu, Chenxin Li, Chen Yang, Yixuan Yuan

Reconstructing deformable tissues from endoscopic videos is essential in many downstream surgical applications. However, existing methods suffer from slow rendering speed, greatly limiting their practical use. In this pa…

3DGSDecoderDepth Estimation