paper-with-me

홈 › Papers

LanGuSTE: Language-Guided Coarse-to-Fine Patch Selection for Efficient Whole Slide Image Analysis

2025-08-20 · Yonghan Shin, SeungKyu Kim, Won-Ki Jeong arxiv

Whole slide images (WSIs) in computational pathology pose a major computational challenge due to their gigapixel scale, often requiring tens to hundreds of thousands of high-resolution patches to be processed per slide. In conventional WSI pipelines, exhaustive high-resolution patch processing makes preprocessing far more time-consuming than downstream model training. Existing patch selection methods suffer from a fundamental paradox: all patches must still be extracted and encoded at least during training, and sometimes during both training and inference, before irrelevant ones can be discarded. To address this, we propose LanGuSTE, an efficient patch selection framework that integrates pathology-domain vision-language models (VLMs) and knowledge derived from large language models (LLMs) through two key modules: Cross- Scale Visual Prompt Tuning (CS-VPT) and coarse-to-fine patch selection. CS-VPT aligns low-resolution patches with their spatially corresponding high-resolution patches through contrastive learning, transferring fine-grained diagnostic semantics into low-resolution representations. The patch selection module then leverages VLM representations and LLM-generated pathology-specific descriptions to identify informative regions in a coarse-to-fine manner, encoding only the corresponding high-resolution patches to reduce preprocessing time. Extensive experiments demonstrate that LanGuSTE reduces overall WSI processing time to approximately 3x while achieving diagnostic performance comparable to or better than exhaustive patch processing and recent state-of-the-art patch-selection methods.

📄 PDF Abstract BibTeX arXiv:2508.14537

Code (0)

등록된 구현이 없습니다.

Tasks

Knowledge Distillation

Similar Papers 제목 키워드 기반

EDGER: EDge-Guided with HEatmap Refinement for Generalizable Image Forgery Localization

2026-05-12 · Minh-Khoa Le-Phan, Minh-Hoang Le, Minh-Triet Tran, Trong-Le Do arxiv

Text-guided inpainting has made image forgery increasingly realistic, challenging both SID and IFL. However, existing methods often struggle to point out suspicious signals across domains. To address this problem, we pro…

Domain Generalization

CoFiNet: Reliable Coarse-to-fine Correspondences for Robust PointCloud Registration

2021-05-21 · NeurIPS 2021 12 · Hao Yu, Fu Li, Mahdi Saleh, Benjamin Busam 외

We study the problem of extracting correspondences between a pair of point clouds for registration. For correspondence retrieval, existing works benefit from matching sparse keypoints detected from dense points but usual…

Keypoint DetectionPoint Cloud RegistrationRetrieval

CoFiNet: Reliable Coarse-to-fine Correspondences for Robust Point Cloud Registration

2021-10-26 · NeurIPS 2021 12 · Hao Yu, Fu Li, Mahdi Saleh, Benjamin Busam 외

We study the problem of extracting correspondences between a pair of point clouds for registration. For correspondence retrieval, existing works benefit from matching sparse keypoints detected from dense points but usual…

Keypoint DetectionPoint Cloud RegistrationRetrieval

Dual-Resolution Correspondence Networks

2020-06-16 · NeurIPS 2020 12 · Xinghui Li, Kai Han, Shuda Li, Victor Adrian Prisacariu

We tackle the problem of establishing dense pixel-wise correspondences between a pair of images. In this work, we introduce Dual-Resolution Correspondence Networks (DualRC-Net), to obtain pixel-wise correspondences in a …

Coarse-to-Fine Learning for Multi-Pipette Localisation in Robot-Assisted In Vivo Patch-Clamp

2025-03-31 · Lan Wei, Gema Vera Gonzalez, Phatsimo Kgwarae, Alexander Timms 외

In vivo image-guided multi-pipette patch-clamp is essential for studying cellular interactions and network dynamics in neuroscience. However, current procedures mainly rely on manual expertise, which limits accessibility…

Generative Adversarial Network