paper-with-me

홈 › Papers

M3ER: Multiplicative Multimodal Emotion Recognition Using Facial, Textual, and Speech Cues

2019-11-09 · Trisha Mittal, Uttaran Bhattacharya, Rohan Chandra, Aniket Bera, Dinesh Manocha

We present M3ER, a learning-based method for emotion recognition from multiple input modalities. Our approach combines cues from multiple co-occurring modalities (such as face, text, and speech) and also is more robust than other methods to sensor noise in any of the individual modalities. M3ER models a novel, data-driven multiplicative fusion method to combine the modalities, which learn to emphasize the more reliable cues and suppress others on a per-sample basis. By introducing a check step which uses Canonical Correlational Analysis to differentiate between ineffective and effective modalities, M3ER is robust to sensor noise. M3ER also generates proxy features in place of the ineffectual modalities. We demonstrate the efficiency of our network through experimentation on two benchmark datasets, IEMOCAP and CMU-MOSEI. We report a mean accuracy of 82.7% on IEMOCAP and 89.0% on CMU-MOSEI, which, collectively, is an improvement of about 5% over prior work.

📄 PDF Abstract BibTeX arXiv:1911.05659

Code (0)

등록된 구현이 없습니다.

Tasks

Emotion RecognitionMultimodal Emotion Recognition

Similar Papers 제목 키워드 기반

MicroEmo: Time-Sensitive Multimodal Emotion Recognition with Micro-Expression Dynamics in Video Dialogues

2024-07-23 · Liyun Zhang

Multimodal Large Language Models (MLLMs) have demonstrated remarkable multimodal emotion recognition capabilities, integrating multimodal cues from visual, acoustic, and linguistic contexts in the video to recognize huma…

Emotion RecognitionMultimodal Emotion Recognition

Salience Adjustment for Context-Based Emotion Recognition

2025-07-17 · Bin Han, Jonathan Gratch arxiv

Emotion recognition in dynamic social contexts requires an understanding of the complex interaction between facial expressions and situational cues. This paper presents a salience-adjusted framework for context-aware emo…

Emotion Recognition

A Facial Expression-Aware Multimodal Multi-task Learning Framework for Emotion Recognition in Multi-party Conversations

2023-07-01 · Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) 2023 7 · Wenjie Zheng, Jianfei Yu, Rui Xia, Shijin Wang

Multimodal Emotion Recognition in Multiparty Conversations (MERMC) has recently attracted considerable attention. Due to the complexity of visual scenes in multi-party conversations, most previous MERMC studies mainly fo…

Emotion RecognitionEmotion Recognition in ConversationFacial Expression Recognition (FER)Multimodal Emotion Recognition+1

Multimodal Emotion Recognition via Bi-directional Cross-Attention and Temporal Modeling

2026-03-12 · Junhyeong Byeon, Jeongyeol Kim, Sejoon Lim arxiv

Expression recognition in in-the-wild video data remains challenging due to substantial variations in facial appearance, background conditions, audio noise, and the inherently dynamic nature of human affect. Relying on a…

Multimodal Emotion RecognitionRepresentation Learning

Textualized and Feature-based Models for Compound Multimodal Emotion Recognition in the Wild

2024-07-17 · Nicolas Richet, Soufiane Belharbi, Haseeb Aslam, Meike Emilie Schadt 외

Systems for multimodal emotion recognition (ER) are commonly trained to extract features from different modalities (e.g., visual, audio, and textual) that are combined to predict individual basic emotions. However, compo…

Emotion RecognitionMultimodal Emotion Recognition