paper-with-me

홈 › Papers

Unified Modeling Enhanced Multimodal Learning for Precision Neuro-Oncology

2024-06-11 · Huahui Yi, Xiaofei Wang, Kang Li, Chao Li

Multimodal learning, integrating histology images and genomics, promises to enhance precision oncology with comprehensive views at microscopic and molecular levels. However, existing methods may not sufficiently model the shared or complementary information for more effective integration. In this study, we introduce a Unified Modeling Enhanced Multimodal Learning (UMEML) framework that employs a hierarchical attention structure to effectively leverage shared and complementary features of both modalities of histology and genomics. Specifically, to mitigate unimodal bias from modality imbalance, we utilize a query-based cross-attention mechanism for prototype clustering in the pathology encoder. Our prototype assignment and modularity strategy are designed to align shared features and minimizes modality gaps. An additional registration mechanism with learnable tokens is introduced to enhance cross-modal feature integration and robustness in multimodal unified modeling. Our experiments demonstrate that our method surpasses previous state-of-the-art approaches in glioma diagnosis and prognosis tasks, underscoring its superiority in precision neuro-Oncology.

📄 PDF Abstract BibTeX arXiv:2406.07078

Code (1)

huahuiyi/mmdp 공식 구현 pytorch

Tasks

Prognosis

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

NeuroVascU-Net: A Unified Multi-Scale and Cross-Domain Adaptive Feature Fusion U-Net for Precise 3D Segmentation of Brain Vessels in Contrast-Enhanced T1 MRI

2025-11-23 · Mohammad Jafari Vayeghan, Niloufar Delfan, Mehdi Tale Masouleh, Mansour Parvaresh Rizi 외 arxiv

Precise 3D segmentation of cerebral vasculature from T1-weighted contrast-enhanced (T1CE) MRI is crucial for safe neurosurgical planning. Manual delineation is time-consuming and prone to inter-observer variability, whil…

BrainFuse: a unified infrastructure integrating realistic biological modeling and core AI methodology

2026-01-29 · Baiyu Chen, Yujie Wu, Siyuan Xu, Peng Qu 외 arxiv

Neuroscience and artificial intelligence represent distinct yet complementary pathways to general intelligence. However, amid the ongoing boom in AI research and applications, the translational synergy between these two …

Latent Representation Learning for Multimodal Brain Activity Translation

2024-09-27 · Arman Afrasiyabi, Dhananjay Bhaskar, Erica L. Busch, Laurent Caplette 외

Neuroscience employs diverse neuroimaging techniques, each offering distinct insights into brain activity, from electrophysiological recordings such as EEG, which have high temporal resolution, to hemodynamic modalities …

EEGFunctional ConnectivityGraph AttentionRepresentation Learning+1

TimeOmni-VL: Unified Models for Time Series Understanding and Generation

2026-02-19 · Tong Guan, Sheng Pan, Johan Barthelemy, Zhao Li 외 arxiv

Recent time series modeling faces a sharp divide between numerical generation and semantic understanding, with research showing that generation models often rely on superficial pattern matching, while understanding-orien…

F2IND-IT! -- Multimodal Fuzzy Fake Indian News Detection using Images and Text

2026-05-16 · Kushal Trivedi, Murtuza Shaikh, Khushi Singh, Jeevaraj S. arxiv

Biased manipulation of facts across regional and national media outlets complicates misinformation detection in diverse landscapes like India. This paper introduces a novel multimodal framework combining visual and textu…

Fake News Detection