paper-with-me

Papers

Probing, Fusion, and Trustworthiness: A Systematic Evaluation of Foundation Model Representations for Multimodal Cancer Analysis

2026-06-15 · Jingyu Hu, Giuseppe Tripodi, Reed Naidoo, Sarah F. McGough, Tapabrata Chakraborti arxiv

Foundation models (FMs) have emerged as powerful representation extractors for medical data, yet their generalizability to datasets under distribution shift remains underexplored. This work systematically evaluates FM-based representations on a suite of computational pathology tasks across two real-world commercial cohorts, IH-BC and IH-NSCLC, drawn from the licensed in-house (IH) oncology dataset. The analysis focuses on two modalities, whole-slide images and transcriptomic profiles, drawn from the IH multimodal data. We first benchmark unimodal probing performance across five FMs on eight downstream classification tasks, and find that image and omics representations carry complementary predictive signals. Then we investigate whether multimodal fusion can yield additional gains over unimodal baselines by comparing three image-omics fusion strategies built on paired representations. The trustworthiness of selected unimodal and multimodal pipelines is further assessed through conformal prediction. Our results show that FM representations achieve competitive performance on out-of-distribution data and that multimodal fusion helps mainly when no single modality dominates the signal. Conformal prediction reveals that in the majority of cases where a point prediction fails, the true diagnosis remains recoverable within the prediction set, reinforcing the value of uncertainty-aware inference for clinical support.

📄 PDF Abstract BibTeX arXiv:2606.17115

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Do You Trust Me? Cognitive-Affective Signatures of Trustworthiness in Large Language Models

2025-12-17 · Gerard Yeo, Svetlana Churina, Kokil Jaidka arxiv

Perceived trustworthiness underpins how users navigate online information, yet it remains unclear whether large language models (LLMs),increasingly embedded in search, recommendation, and conversational systems, represen…

Do Transformers Encode a Foundational Ontology? Probing Abstract Classes in Natural Language

2022-01-25 · Mael Jullien, Marco Valentino, Andre Freitas

With the methodological support of probing (or diagnostic classification), recent studies have demonstrated that Transformers encode syntactic and semantic information to some extent. Following this line of research, thi…

Diagnostic

VMDT: Decoding the Trustworthiness of Video Foundation Models

2025-11-07 · Yujin Potter, Zhun Wang, Nicholas Crispino, Kyle Montgomery 외 arxiv

As foundation models become more sophisticated, ensuring their trustworthiness becomes increasingly critical; yet, unlike text and image, the video modality still lacks comprehensive trustworthiness benchmarks. We introd…

Adversarial Robustness

TrustLDM: Benchmarking Trustworthiness in Language Diffusion Models

2026-04-15 · Yichuan Mo, Yukun Jiang, Yanbo Shi, Mingjie Li 외 arxiv

The rapid development of Language Diffusion Models (LDMs) challenges the dominant position of auto-regressive competitors in language processing. However, their flexible, any-order decoding strategies not only enable fas…

Towards Tracing Trustworthiness Dynamics: Revisiting Pre-training Period of Large Language Models

2024-02-29 · Chen Qian, Jie Zhang, Wei Yao, Dongrui Liu 외

Ensuring the trustworthiness of large language models (LLMs) is crucial. Most studies concentrate on fully pre-trained LLMs to better understand and improve LLMs' trustworthiness. In this paper, to reveal the untapped po…

FairnessMutual Information Estimation