paper-with-me

Papers

MMSF: Multitask and Multimodal Supervised Framework for WSI Classification and Survival Analysis

2026-01-28 · Chengying She, Chengwei Chen, Xinran Zhang, Ben Wang, Lizhuang Liu, Chengwei Shao, Yun Bian arxiv

Multimodal evidence is critical in computational pathology: gigapixel whole slide images capture tumor morphology, while patient-level clinical descriptors preserve complementary context for prognosis. Integrating such heterogeneous signals remains challenging because feature spaces exhibit distinct statistics and scales. We introduce MMSF, a multitask and multimodal supervised framework built on a linear-complexity MIL backbone that explicitly decomposes and fuses cross-modal information. MMSF comprises a graph feature extraction module embedding tissue topology at the patch level, a clinical data embedding module standardizing patient attributes, a feature fusion module aligning modality-shared and modality-specific representations, and a Mamba-based MIL encoder with multitask prediction heads. Experiments on CAMELYON16 and TCGA-NSCLC demonstrate 2.1--6.6\% accuracy and 2.2--6.9\% AUC improvements over competitive baselines, while evaluations on five TCGA survival cohorts yield 7.1--9.8\% C-index improvements compared with unimodal methods and 5.6--7.1\% over multimodal alternatives.

📄 PDF Abstract BibTeX arXiv:2601.20347

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

MMSFormer: Multimodal Transformer for Material and Semantic Segmentation

2023-09-07 · Md Kaykobad Reza, Ashley Prater-Bennette, M. Salman Asif

Leveraging information across diverse modalities is known to enhance performance on multimodal segmentation tasks. However, effectively fusing information from different modalities remains challenging due to the unique c…

SegmentationSemantic SegmentationThermal Image Segmentation

Weakly Supervised Multimodal Temporal Forgery Localization via Multitask Learning

2025-08-04 · Wenbo Xu, Wei Lu, Xiangyang Luo arxiv

The spread of Deepfake videos has caused a trust crisis and impaired social stability. Although numerous approaches have been proposed to address the challenges of Deepfake detection and localization, there is still a la…

Binary ClassificationDeepFake Detection

Fusion for Visual-Infrared Person ReID in Real-World Surveillance Using Corrupted Multimodal Data

2023-04-29 · Arthur Josi, Mahdi Alehdaghi, Rafael M. O. Cruz, Eric Granger

Visible-infrared person re-identification (V-I ReID) seeks to match images of individuals captured over a distributed network of RGB and IR cameras. The task is challenging due to the significant differences between V an…

Data AugmentationPerson Re-Identification

M3H: Multimodal Multitask Machine Learning for Healthcare

2024-04-29 · Dimitris Bertsimas, Yu Ma

Developing an integrated many-to-many framework leveraging multimodal data for multiple tasks is crucial to unifying healthcare applications ranging from diagnoses to operations. In resource-constrained hospital environm…

Binary ClassificationPatient PhenotypingTime Series

M&M: Multimodal-Multitask Model Integrating Audiovisual Cues in Cognitive Load Assessment

2024-03-14 · Long Nguyen-Phuoc, Renald Gaboriau, Dimitri Delacroix, Laurent Navarro

This paper introduces the M&M model, a novel multimodal-multitask learning framework, applied to the AVCAffe dataset for cognitive load assessment (CLA). M&M uniquely integrates audiovisual cues through a dual-pathway ar…