paper-with-me

Papers

Cross-modality Attention-based Multimodal Fusion for Non-small Cell Lung Cancer (NSCLC) Patient Survival Prediction

2023-08-18 · Ruining Deng, Nazim Shaikh, Gareth Shannon, Yao Nie

Cancer prognosis and survival outcome predictions are crucial for therapeutic response estimation and for stratifying patients into various treatment groups. Medical domains concerned with cancer prognosis are abundant with multiple modalities, including pathological image data and non-image data such as genomic information. To date, multimodal learning has shown potential to enhance clinical prediction model performance by extracting and aggregating information from different modalities of the same subject. This approach could outperform single modality learning, thus improving computer-aided diagnosis and prognosis in numerous medical applications. In this work, we propose a cross-modality attention-based multimodal fusion pipeline designed to integrate modality-specific knowledge for patient survival prediction in non-small cell lung cancer (NSCLC). Instead of merely concatenating or summing up the features from different modalities, our method gauges the importance of each modality for feature fusion with cross-modality relationship when infusing the multimodal features. Compared with single modality, which achieved c-index of 0.5772 and 0.5885 using solely tissue image data or RNA-seq data, respectively, the proposed fusion approach achieved c-index 0.6587 in our experiment, showcasing the capability of assimilating modality-specific knowledge from varied modalities.

📄 PDF Abstract BibTeX arXiv:2308.09831

Code (1)

hrlblab/cs-mil pytorch

Tasks

PrognosisSurvival Prediction

Similar Papers 제목 키워드 기반

MANGO: Multimodal Attention-based Normalizing Flow Approach to Fusion Learning

2025-08-13 · Thanh-Dat Truong, Christophe Bobda, Nitin Agarwal, Khoa Luu arxiv

Multimodal learning has gained much success in recent years. However, current multimodal fusion methods adopt the attention mechanism of Transformers to implicitly learn the underlying correlation of multimodal features.…

Image-to-Image TranslationSemantic SegmentationGenre classification

DeepMLF: Multimodal language model with learnable tokens for deep fusion in sentiment analysis

2025-04-15 · Efthymios Georgiou, Vassilis Katsouros, Yannis Avrithis, Alexandros Potamianos

While multimodal fusion has been extensively studied in Multimodal Sentiment Analysis (MSA), the role of fusion depth and multimodal capacity allocation remains underexplored. In this work, we position fusion depth, scal…

DecoderLanguage ModelingLanguage ModellingMultimodal Sentiment Analysis+1

MedMIX: Modality-Internal Expert Fusion for Multimodal Medical Diagnosis

2026-05-15 · Seungik Cho, Anqi Li, Wei Qiu arxiv

Multimodal clinical prediction faces three challenges: multiple foundation models (FMs) with complementary strengths per modality, pervasive missing modalities at training and test time, and sample-specific variation in …

Medical Diagnosis

Exploring Attention Mechanisms for Multimodal Emotion Recognition in an Emergency Call Center Corpus

2023-06-12 · Théo Deschamps-Berger, Lori Lamel, Laurence Devillers

The emotion detection technology to enhance human decision-making is an important research issue for real-world applications, but real-life emotion datasets are relatively rare and small. The experiments conducted in thi…

Decision MakingEmotion RecognitionMultimodal Emotion RecognitionSpeech Emotion Recognition

Asynchronous Multimodal Video Sequence Fusion via Learning Modality-Exclusive and -Agnostic Representations

2024-07-06 · Dingkang Yang, Mingcheng Li, Linhao Qu, Kun Yang 외

Understanding human intentions (e.g., emotions) from videos has received considerable attention recently. Video streams generally constitute a blend of temporal data stemming from distinct modalities, including natural l…