paper-with-me

Papers

Image Retrieval with Intra-Sweep Representation Learning for Neck Ultrasound Scanning Guidance

2024-12-10 · Wanwen Chen, Adam Schmidt, Eitan Prisman, Septimiu E. Salcudean

Purpose: Intraoperative ultrasound (US) can enhance real-time visualization in transoral robotic surgery. The surgeon creates a mental map with a pre-operative scan. Then, a surgical assistant performs freehand US scanning during the surgery while the surgeon operates at the remote surgical console. Communicating the target scanning plane in the surgeon's mental map is difficult. Automatic image retrieval can help match intraoperative images to preoperative scans, guiding the assistant to adjust the US probe toward the target plane. Methods: We propose a self-supervised contrastive learning approach to match intraoperative US views to a preoperative image database. We introduce a novel contrastive learning strategy that leverages intra-sweep similarity and US probe location to improve feature encoding. Additionally, our model incorporates a flexible threshold to reject unsatisfactory matches. Results: Our method achieves 92.30% retrieval accuracy on simulated data and outperforms state-of-the-art temporal-based contrastive learning approaches. Our ablation study demonstrates that using probe location in the optimization goal improves image representation, suggesting that semantic information can be extracted from probe location. We also present our approach on real patient data to show the feasibility of the proposed US probe localization system despite tissue deformation from tongue retraction. Conclusion: Our contrastive learning method, which utilizes intra-sweep similarity and US probe location, enhances US image representation learning. We also demonstrate the feasibility of using our image retrieval method to provide neck US localization on real patient US after tongue retraction.

📄 PDF Abstract BibTeX arXiv:2412.07741

Code (0)

등록된 구현이 없습니다.

Tasks

Contrastive LearningImage RetrievalRepresentation LearningRetrieval

Methods 이 논문이 사용한 방법론

Contrastive Learning 설명 없음

Similar Papers 제목 키워드 기반

Intra-Modal Constraint Loss For Image-Text Retrieval

2022-07-11 · Jianan Chen, Lu Zhang, Qiong Wang, Cong Bai 외

Cross-modal retrieval has drawn much attention in both computer vision and natural language processing domains. With the development of convolutional and recurrent neural networks, the bottleneck of retrieval across imag…

Cross-Modal RetrievalImage-text RetrievalRetrievalText Retrieval

LexLIP: Lexicon-Bottlenecked Language-Image Pre-Training for Large-Scale Image-Text Retrieval

2023-02-06 · Ziyang Luo, Pu Zhao, Can Xu, Xiubo Geng 외

Image-text retrieval (ITR) is a task to retrieve the relevant images/texts, given the query from another modality. The conventional dense retrieval paradigm relies on encoding images and texts into dense representations …

Image-text RetrievalRetrievalText Retrieval

LexLIP: Lexicon-Bottlenecked Language-Image Pre-Training for Large-Scale Image-Text Sparse Retrieval

2023-01-01 · ICCV 2023 1 · Ziyang Luo, Pu Zhao, Can Xu, Xiubo Geng 외

Image-text retrieval (ITR) aims to retrieve images or texts that match a query originating from the other modality. The conventional dense retrieval paradigm relies on encoding images and texts into dense representat…

image-classificationImage ClassificationImage-text RetrievalRetrieval+2

Which Speech Representation Better Matches Text-Native Reasoning? A Study of Speech-Text Alignment on Frame Rate and Representation

2026-06-10 · Zhen Ye, Xu Tan, Yiming Li, Guangyan Zhang 외 arxiv

Spoken dialogue models typically start from text LLM backbones, yet reasoning often degrades when conditioning on speech instead of text. We attribute part of this modality gap to a temporal-granularity mismatch: speech …

FS-I2P:A Hierarchical Focus-Sweep Registration Network with Dynamically Allocated Depth

2026-05-08 · Zhixin Cheng, Yujia Chen, Xujing Tao, Bohao Liao 외 arxiv

Image-to-point cloud registration is often challenged by viewpoint changes, cross-modal discrepancies, and repetitive textures, which induce scale ambiguity and consequently lead to erroneous correspondences. Recent dete…

Point Cloud Registration