paper-with-me

Papers

Flex-MoE: Modeling Arbitrary Modality Combination via the Flexible Mixture-of-Experts

2024-10-10 · Sukwon Yun, Inyoung Choi, Jie Peng, Yangfan Wu, Jingxuan Bao, Qiyiwen Zhang, Jiayi Xin, Qi Long, Tianlong Chen

Multimodal learning has gained increasing importance across various fields, offering the ability to integrate data from diverse sources such as images, text, and personalized records, which are frequently observed in medical domains. However, in scenarios where some modalities are missing, many existing frameworks struggle to accommodate arbitrary modality combinations, often relying heavily on a single modality or complete data. This oversight of potential modality combinations limits their applicability in real-world situations. To address this challenge, we propose Flex-MoE (Flexible Mixture-of-Experts), a new framework designed to flexibly incorporate arbitrary modality combinations while maintaining robustness to missing data. The core idea of Flex-MoE is to first address missing modalities using a new missing modality bank that integrates observed modality combinations with the corresponding missing ones. This is followed by a uniquely designed Sparse MoE framework. Specifically, Flex-MoE first trains experts using samples with all modalities to inject generalized knowledge through the generalized router ($\mathcal{G}$-Router). The $\mathcal{S}$-Router then specializes in handling fewer modality combinations by assigning the top-1 gate to the expert corresponding to the observed modality combination. We evaluate Flex-MoE on the ADNI dataset, which encompasses four modalities in the Alzheimer's Disease domain, as well as on the MIMIC-IV dataset. The results demonstrate the effectiveness of Flex-MoE highlighting its ability to model arbitrary modality combinations in diverse missing modality scenarios. Code is available at https://github.com/UNITES-Lab/flex-moe.

📄 PDF Abstract BibTeX arXiv:2410.08245

Code (1)

unites-lab/flex-moe 공식 구현 pytorch

Tasks

Mixture-of-Experts

Methods 이 논문이 사용한 방법론

MoE 설명 없음

Similar Papers 제목 키워드 기반

AnyMo: Scaling Any-Modality Conditional Motion Generation with Masked Modeling

2026-05-28 · Yiheng Li, Zhuo Li, Ruibing Hou, Yingjie Chen 외 arxiv

Conditional human motion generation remains a fundamental challenge in computer vision and robotics. Despite significant progress, current methods are often constrained by fixed modality configurations and task-specific …

Motion Synthesis

FlexiReID: Adaptive Mixture of Expert for Multi-Modal Person Re-Identification

2025-10-17 · Zhen Sun, Lei Tan, Yunhang Shen, Chengmao Cai 외 arxiv

Multimodal person re-identification (Re-ID) aims to match pedestrian images across different modalities. However, most existing methods focus on limited cross-modal settings and fail to support arbitrary query-retrieval …

Person Re-Identification

OmniSegmentor: A Flexible Multi-Modal Learning Framework for Semantic Segmentation

2025-09-18 · Bo-Wen Yin, Jiao-Long Cao, Xuying Zhang, Yuming Chen 외 arxiv

Recent research on representation learning has proved the merits of multi-modal clues for robust semantic segmentation. Nevertheless, a flexible pretrain-and-finetune pipeline for multiple visual modalities remains unexp…

Representation LearningSemantic Segmentation

OpenDance: Multimodal Controllable 3D Dance Generation Using Large-scale Internet Data

2025-06-09 · Jinlu Zhang, Zixi Kang, Yizhou Wang

Music-driven dance generation offers significant creative potential yet faces considerable challenges. The absence of fine-grained multimodal data and the difficulty of flexible multi-conditional generation limit previou…

Diversity

Flexible Fusion Network for Multi-modal Brain Tumor Segmentation

2023-05-01 · journal 2023 5 · Hengyi Yang; Tao Zhou; Yi Zhou; Yizhe Zhang; Huazhu Fu

Automated brain tumor segmentation is crucial for aiding brain disease diagnosis and evaluating disease progress. Currently, magnetic resonance imaging (MRI) is a routinely adopted approach in the field of brain tumor se…

Brain Tumor SegmentationDecoderSegmentationTumor Segmentation