paper-with-me

홈 › Papers

Beyond Simple Fusion: Adaptive Gated Fusion for Robust Multimodal Sentiment Analysis

2025-10-02 · Han Wu, Yanming Sun, Yunhe Yang, Derek F. Wong arxiv

Multimodal sentiment analysis (MSA) leverages information fusion from diverse modalities (e.g., text, audio, visual) to enhance sentiment prediction. However, simple fusion techniques often fail to account for variations in modality quality, such as those that are noisy, missing, or semantically conflicting. This oversight leads to suboptimal performance, especially in discerning subtle emotional nuances. To mitigate this limitation, we introduce a simple yet efficient \textbf{A}daptive \textbf{G}ated \textbf{F}usion \textbf{N}etwork that adaptively adjusts feature weights via a dual gate fusion mechanism based on information entropy and modality importance. This mechanism mitigates the influence of noisy modalities and prioritizes informative cues following unimodal encoding and cross-modal interaction. Experiments on CMU-MOSI and CMU-MOSEI show that AGFN significantly outperforms strong baselines in accuracy, effectively discerning subtle emotions with robust performance. Visualization analysis of feature representations demonstrates that AGFN enhances generalization by learning from a broader feature distribution, achieved by reducing the correlation between feature location and prediction error, thereby decreasing reliance on specific locations and creating more robust multimodal feature representations.

📄 PDF Abstract BibTeX arXiv:2510.01677

Code (0)

등록된 구현이 없습니다.

Tasks

Multimodal Sentiment Analysis

Similar Papers 제목 키워드 기반

LiRaFusion: Deep Adaptive LiDAR-Radar Fusion for 3D Object Detection

2024-02-18 · Jingyu Song, Lingjun Zhao, Katherine A. Skinner

We propose LiRaFusion to tackle LiDAR-radar fusion for 3D object detection to fill the performance gap of existing LiDAR-radar detectors. To improve the feature extraction capabilities from these two modalities, we desig…

3D Object Detectionobject-detectionObject Detection

AG-Fusion: adaptive gated multimodal fusion for 3d object detection in complex scenes

2025-10-27 · Sixian Liu, Chen Xu, Qiang Wang, Donghai Shi 외 arxiv

Multimodal camera-LiDAR fusion technology has found extensive application in 3D object detection, demonstrating encouraging performance. However, existing methods exhibit significant performance degradation in challengin…

3D Object Detection

DiffFNO: Diffusion Fourier Neural Operator

2024-11-15 · CVPR 2025 1 · Xiaoyi Liu, Hao Tang

We introduce DiffFNO, a novel diffusion framework for arbitrary-scale super-resolution strengthened by a Weighted Fourier Neural Operator (WFNO). Mode Re-balancing in WFNO effectively captures critical frequency componen…

Computational EfficiencySuper-Resolution

Data-Efficient Ensemble Weather Forecasting with Diffusion Models

2025-09-14 · Kevin Valencia, Ziyang Liu, Justin Cui arxiv

Although numerical weather forecasting methods have dominated the field, recent advances in deep learning methods, such as diffusion models, have shown promise in ensemble weather forecasting. However, such models are ty…

Weather Forecasting

3D Gated Recurrent Fusion for Semantic Scene Completion

2020-02-17 · Yu Liu, Jie Li, Qingsen Yan, Xia Yuan 외

This paper tackles the problem of data fusion in the semantic scene completion (SSC) task, which can simultaneously deal with semantic labeling and scene completion. RGB images contain texture details of the object(s) wh…

3D Semantic Scene CompletionScene Understanding