paper-with-me

Papers

Prototype-Guided Cross-Modal Knowledge Enhancement for Adaptive Survival Prediction

2025-03-13 · Fengchun Liu, Linghan Cai, Zhikang Wang, Zhiyuan Fan, Jin-Gang Yu, Hao Chen, Yongbing Zhang

Histo-genomic multimodal survival prediction has garnered growing attention for its remarkable model performance and potential contributions to precision medicine. However, a significant challenge in clinical practice arises when only unimodal data is available, limiting the usability of these advanced multimodal methods. To address this issue, this study proposes a prototype-guided cross-modal knowledge enhancement (ProSurv) framework, which eliminates the dependency on paired data and enables robust learning and adaptive survival prediction. Specifically, we first introduce an intra-modal updating mechanism to construct modality-specific prototype banks that encapsulate the statistics of the whole training set and preserve the modality-specific risk-relevant features/prototypes across intervals. Subsequently, the proposed cross-modal translation module utilizes the learned prototypes to enhance knowledge representation for multimodal inputs and generate features for missing modalities, ensuring robust and adaptive survival prediction across diverse scenarios. Extensive experiments on four public datasets demonstrate the superiority of ProSurv over state-of-the-art methods using either unimodal or multimodal input, and the ablation study underscores its feasibility for broad applicability. Overall, this study addresses a critical practical challenge in computational pathology, offering substantial significance and potential impact in the field.

📄 PDF Abstract BibTeX arXiv:2503.10726

Code (1)

cyclexfy/ProSurv 공식 구현 pytorch

Tasks

Survival Prediction

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음
SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

SPENet: Self-guided Prototype Enhancement Network for Few-shot Medical Image Segmentation

2025-09-03 · Chao Fan, Xibin Jia, Anqi Xiao, Hongyuan Yu 외 arxiv

Few-Shot Medical Image Segmentation (FSMIS) aims to segment novel classes of medical objects using only a few labeled images. Prototype-based methods have made significant progress in addressing FSMIS. However, they typi…

Medical Image Segmentation

Aurora: Towards Universal Generative Multimodal Time Series Forecasting

2025-09-26 · Xingjian Wu, Jianxin Jin, Wanghui Qiu, Peng Chen 외 arxiv

Cross-domain generalization is very important in Time Series Forecasting because similar historical information may lead to distinct future trends due to the domain-specific characteristics. Recent works focus on buildin…

Time Series ForecastingDomain Generalization

Identity Clue Refinement and Enhancement for Visible-Infrared Person Re-Identification

2025-12-04 · Guoqing Zhang, Zhun Wang, Hairui Wang, Zhonglin Ye 외 arxiv

Visible-Infrared Person Re-Identification (VI-ReID) is a challenging cross-modal matching task due to significant modality discrepancies. While current methods mainly focus on learning modality-invariant features through…

Person Re-Identification

Correlation-Decoupled Knowledge Distillation for Multimodal Sentiment Analysis with Incomplete Modalities

2024-04-25 · CVPR 2024 1 · Mingcheng Li, Dingkang Yang, Xiao Zhao, Shuaibing Wang 외

Multimodal sentiment analysis (MSA) aims to understand human sentiment through multimodal data. Most MSA efforts are based on the assumption of modality completeness. However, in real-world applications, some practical f…

DisentanglementKnowledge DistillationMultimodal Sentiment AnalysisSentiment Analysis

ViewRefer: Grasp the Multi-view Knowledge for 3D Visual Grounding with GPT and Prototype Guidance

2023-03-29 · Zoey Guo, Yiwen Tang, Ray Zhang, Dong Wang 외

Understanding 3D scenes from multi-view inputs has been proven to alleviate the view discrepancy issue in 3D visual grounding. However, existing methods normally neglect the view cues embedded in the text modality and fa…

3D visual groundingVisual Grounding