paper-with-me

홈 › Papers

Cross-modal Prompting for Balanced Incomplete Multi-modal Emotion Recognition

2025-12-12 · Wen-Jue He, Xiaofeng Zhu, Zheng Zhang arxiv

Incomplete multi-modal emotion recognition (IMER) aims at understanding human intentions and sentiments by comprehensively exploring the partially observed multi-source data. Although the multi-modal data is expected to provide more abundant information, the performance gap and modality under-optimization problem hinder effective multi-modal learning in practice, and are exacerbated in the confrontation of the missing data. To address this issue, we devise a novel Cross-modal Prompting (ComP) method, which emphasizes coherent information by enhancing modality-specific features and improves the overall recognition accuracy by boosting each modality's performance. Specifically, a progressive prompt generation module with a dynamic gradient modulator is proposed to produce concise and consistent modality semantic cues. Meanwhile, cross-modal knowledge propagation selectively amplifies the consistent information in modality features with the delivered prompts to enhance the discrimination of the modality-specific output. Additionally, a coordinator is designed to dynamically re-weight the modality outputs as a complement to the balance strategy to improve the model's efficacy. Extensive experiments on 4 datasets with 7 SOTA methods under different missing rates validate the effectiveness of our proposed method.

📄 PDF Abstract BibTeX arXiv:2512.11239

Code (0)

등록된 구현이 없습니다.

Tasks

Emotion Recognition

Similar Papers 제목 키워드 기반

PASSION: Towards Effective Incomplete Multi-Modal Medical Image Segmentation with Imbalanced Missing Rates

2024-07-20 · Junjie Shi, Caozhi Shang, Zhaobin Sun, Li Yu 외

Incomplete multi-modal image segmentation is a fundamental task in medical imaging to refine deployment efficiency when only partial modalities are available. However, the common practice that complete-modality data is v…

Image SegmentationMedical Image SegmentationMulti-modal image segmentationSemantic Segmentation

SGMA: Semantic-Guided Modality-Aware Segmentation for Remote Sensing with Incomplete Multimodal Data

2026-03-03 · Lekang Wen, Liang Liao, Jing Xiao, Mi Wang arxiv

Multimodal semantic segmentation integrates complementary information from diverse sensors for remote sensing Earth observation. However, practical systems often encounter missing modalities due to sensor failures or inc…

Semantic SegmentationContrastive Learning

Consistency-Aware Padding for Incomplete Multi-Modal Alignment Clustering Based on Self-Repellent Greedy Anchor Search

2025-07-05 · Shubin Ma, Liang Zhao, Mingdong Lu, Yifan Guo 외 arxiv

Multimodal representation is faithful and highly effective in describing real-world data samples' characteristics by describing their complementary information. However, the collected data often exhibits incomplete and m…

Contrastive Learning

Distilled Prompt Learning for Incomplete Multimodal Survival Prediction

2025-01-01 · CVPR 2025 1 · Yingxue Xu, Fengtao Zhou, Chenyu Zhao, Yihui Wang 외

The integration of multimodal data including pathology images and gene profiles is widely applied in precise survival prediction. Despite recent advances in multimodal survival models, collecting complete modalities …

PredictionPrompt LearningSurvival Prediction

MissBench: Benchmarking Multimodal Affective Analysis under Imbalanced Missing Modalities

2026-03-10 · Tien Anh Pham, Phuong-Anh Nguyen, Duc-Trong Le, Cam-Van Thi Nguyen arxiv

Multimodal affective computing underpins key tasks such as sentiment analysis and emotion recognition. Standard evaluations, however, often assume that textual, acoustic, and visual modalities are equally available. In r…

Emotion RecognitionSentiment Analysis