paper-with-me

홈 › Papers

Self-supervision drives representational convergence in medical foundation models more than clinical supervision

2026-07-22 · Soroosh Tayebi Arasteh, Sebastian Ziegelmayer, Mahshad Lotfinia, Lisa Adams, Sven Nebelung, Jakob Nikolas Kather, Daniel Truhn arxiv

Medical image encoders from different groups are increasingly treated as interchangeable, on the assumption that scale and clinical supervision concentrate their representations onto a shared structure. Whether this convergence is real, what produces it, and whether it is clinically usable are untested, and the similarity measures behind such claims are fragile. We present a controlled dissection across 18 image and 7 text encoders, all open-weight and run locally, spanning 7M to 27B parameters and five imaging modalities, including 650,982 chest radiographs from six datasets. To isolate cause, we train encoders that vary only the objective under fixed data, architecture, and scale, and reproduce the effect in a synthetic model. Convergence is modest but above a random floor, driven by the self-supervised objective, not clinical supervision: matched self-supervised encoders aligned most (40.4% on chest radiography), with label-supervised (21.1%) and image-text (3.3%) far lower, and did not grow with size (Spearman 0.302, p=0.223) or capability. It is within-modality, does not reach clinical language, and does not reproduce how radiologists judge case similarity. Yet a linear classifier transfers across encoders and to five held-out hospitals, retaining about 85% of within-encoder performance. Convergence in medical imaging is therefore set by the pretraining objective, not inherited from scale or clinical supervision. Interoperability is accordingly something to design for through that objective, and to validate where the shared geometry is weakest, across patient subgroups and against clinical judgment.

📄 PDF Abstract BibTeX arXiv:2607.20274

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Training objective drives the consistency of representational similarity across datasets

2024-11-08 · Laure Ciernik, Lorenz Linhardt, Marco Morik, Jonas Dippel 외

The Platonic Representation Hypothesis claims that recent foundation models are converging to a shared representation space as a function of their downstream task performance, irrespective of the objectives and data moda…

image-classificationImage Classification

RelativeFlow: Taming Medical Image Denoising Learning with Noisy Reference

2026-04-16 · Yuxin Liu, Yiqing Dong, Wenxue Yu, Zhan Wu 외 arxiv

Medical image denoising (MID) lacks absolutely clean images for supervision, leading to a noisy reference problem that fundamentally limits denoising performance. Existing simulated-supervised discriminative learning (Si…

Self-Supervised LearningMedical Image Denoising

DiSSECT: Structuring Transfer-Ready Medical Image Representations through Discrete Self-Supervision

2025-09-23 · Azad Singh, Deepak Mishra arxiv

Self-supervised learning (SSL) has emerged as a powerful paradigm for medical image representation learning, particularly in settings with limited labeled data. However, existing SSL methods often rely on complex archite…

Self-Supervised LearningRepresentation Learning

CHERRY: Compressed Hierarchical Experts with Recurrent Representational Yield

2026-06-30 · Dohyeon Kwon, Youngjin Park arxiv

Frontier language capability is usually bought with frontier compute; CHERRY shows a different trade. It is a sovereign Korean model family built on one principle: supervise the tokens that decide the answer, and let sha…

Self-Supervised Flow Matching for Scalable Multi-Modal Synthesis

2026-03-06 · Hila Chefer, Patrick Esser, Dominik Lorenz, Dustin Podell 외 arxiv

Strong semantic representations improve the convergence and generation quality of diffusion and flow models. Existing approaches largely rely on external models, which require separate training, operate on misaligned obj…

Representation LearningAudio Generation