paper-with-me

홈 › Papers

THUNDER: Tile-level Histopathology image UNDERstanding benchmark

2025-07-10 · Pierre Marza, Leo Fillioux, Sofiène Boutaj, Kunal Mahatha, Christian Desrosiers, Pablo Piantanida, Jose Dolz, Stergios Christodoulidis, Maria Vakalopoulou arxiv

Progress in a research field can be hard to assess, in particular when many concurrent methods are proposed in a short period of time. This is the case in digital pathology, where many foundation models have been released recently to serve as feature extractors for tile-level images, being used in a variety of downstream tasks, both for tile- and slide-level problems. Benchmarking available methods then becomes paramount to get a clearer view of the research landscape. In particular, in critical domains such as healthcare, a benchmark should not only focus on evaluating downstream performance, but also provide insights about the main differences between methods, and importantly, further consider uncertainty and robustness to ensure a reliable usage of proposed models. For these reasons, we introduce THUNDER, a tile-level benchmark for digital pathology foundation models, allowing for efficient comparison of many models on diverse datasets with a series of downstream tasks, studying their feature spaces and assessing the robustness and uncertainty of predictions informed by their embeddings. THUNDER is a fast, easy-to-use, dynamic benchmark that can already support a large variety of state-of-the-art foundation, as well as local user-defined models for direct tile-based comparison. In this paper, we provide a comprehensive comparison of 23 foundation models on 16 different datasets covering diverse tasks, feature analysis, and robustness. The code for THUNDER is publicly available at https://github.com/MICS-Lab/thunder.

📄 PDF Abstract BibTeX arXiv:2507.07860

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

TICON: A Slide-Level Tile Contextualizer for Histopathology Representation Learning

2025-12-24 · Varun Belagali, Saarthak Kapse, Pierre Marza, Srijan Das 외 arxiv

The interpretation of small tiles in large whole slide images (WSI) often needs a larger image context. We introduce TICON, a transformer-based tile representation contextualizer that produces rich, contextualized embedd…

Representation Learning

Thunder-KoNUBench: A Corpus-Aligned Benchmark for Korean Negation Understanding

2026-01-08 · Sungmok Jung, Yeonkyoung So, Joonhak Lee, Sangho Kim 외 arxiv

Although negation is known to challenge large language models (LLMs), benchmarks for evaluating negation understanding-especially in Korean-are scarce. We conduct a corpus-based analysis of Korean negation and show that …

Thunder-NUBench: A Benchmark for LLMs' Sentence-Level Negation Understanding

2025-06-17 · Yeonkyoung So, Gyuseong Lee, Sungmok Jung, Joonhak Lee 외

Negation is a fundamental linguistic phenomenon that poses persistent challenges for Large Language Models (LLMs), particularly in tasks requiring deep semantic understanding. Existing benchmarks often treat negation as …

Multiple-choiceNatural Language InferenceNegationSentence

Deep Blur Multi-Model (DeepBlurMM) -- a strategy to mitigate the impact of image blur on deep learning model performance in histopathology image analysis

2024-05-15 · Yujie Xiang, Bojing Liu, Mattias Rantalainen

AI-based analysis of histopathology whole slide images (WSIs) is central in computational pathology. However, image quality, including unsharp areas of WSIs, impacts model performance. We investigate the impact of blur a…

Binary Classificationwhole slide images

DISC: Latent Diffusion Models with Self-Distillation from Separated Conditions for Prostate Cancer Grading

2024-04-19 · Man M. Ho, Elham Ghelichkhan, Yosep Chong, Yufei Zhou 외

Latent Diffusion Models (LDMs) can generate high-fidelity images from noise, offering a promising approach for augmenting histopathology images for training cancer grading models. While previous works successfully genera…