paper-with-me

홈 › Papers

Joint Multimodal Transformer for Emotion Recognition in the Wild

2024-03-15 · Paul Waligora, Haseeb Aslam, Osama Zeeshan, Soufiane Belharbi, Alessandro Lameiras Koerich, Marco Pedersoli, Simon Bacon, Eric Granger

Multimodal emotion recognition (MMER) systems typically outperform unimodal systems by leveraging the inter- and intra-modal relationships between, e.g., visual, textual, physiological, and auditory modalities. This paper proposes an MMER method that relies on a joint multimodal transformer (JMT) for fusion with key-based cross-attention. This framework can exploit the complementary nature of diverse modalities to improve predictive accuracy. Separate backbones capture intra-modal spatiotemporal dependencies within each modality over video sequences. Subsequently, our JMT fusion architecture integrates the individual modality embeddings, allowing the model to effectively capture inter- and intra-modal relationships. Extensive experiments on two challenging expression recognition tasks -- (1) dimensional emotion recognition on the Affwild2 dataset (with face and voice) and (2) pain estimation on the Biovid dataset (with face and biosensors) -- indicate that our JMT fusion can provide a cost-effective solution for MMER. Empirical results show that MMER systems with our proposed fusion allow us to outperform relevant baseline and state-of-the-art methods.

📄 PDF Abstract BibTeX arXiv:2403.10488

Code (1)

PoloWlg/Joint-Multimodal-Transformer-6th-ABAW 공식 구현 pytorch

Tasks

Emotion RecognitionMultimodal Emotion Recognition

Similar Papers 제목 키워드 기반

Feature-Based Dual Visual Feature Extraction Model for Compound Multimodal Emotion Recognition

2025-03-21 · Ran Liu, Fengyu Zhang, Cong Yu, Longjiang Yang 외

This article presents our results for the eighth Affective Behavior Analysis in-the-wild (ABAW) competition.Multimodal emotion recognition (ER) has important applications in affective computing and human-computer interac…

Emotion RecognitionMultimodal Emotion Recognition

A Transformer-based joint-encoding for Emotion Recognition and Sentiment Analysis

2020-06-29 · WS 2020 7 · Jean-Benoit Delbrouck, Noé Tits, Mathilde Brousmiche, Stéphane Dupont

Understanding expressed sentiment and emotions are two crucial factors in human multimodal language. This paper describes a Transformer-based joint-encoding (TBJE) for the task of Emotion Recognition and Sentiment Analys…

Emotion RecognitionMultimodal Sentiment AnalysisSentiment Analysis

Solution to the 10th ABAW Expression Recognition Challenge: A Robust Multimodal Framework with Safe Cross-Attention and Modality Dropout

2026-03-09 · Jun Yu, Naixiang Zheng, Guoyuan Wang, Yunxiang Zhang 외 arxiv

Emotion recognition in real-world environments is hindered by partial occlusions, missing modalities, and severe class imbalance. To address these issues, particularly for the Affective Behavior Analysis in-the-wild (ABA…

Emotion Recognition

Multimodal Group Emotion Recognition In-the-wild Using Privacy-Compliant Features

2023-12-06 · Anderson Augusma, Dominique Vaufreydaz, Frédérique Letué

This paper explores privacy-compliant group-level emotion recognition ''in-the-wild'' within the EmotiW Challenge 2023. Group-level emotion recognition can be useful in many fields including social robotics, conversation…

Emotion Recognition

EmotiW 2018: Audio-Video, Student Engagement and Group-Level Affect Prediction

2018-08-23 · Abhinav Dhall, Amanjot Kaur, Roland Goecke, Tom Gedeon

This paper details the sixth Emotion Recognition in the Wild (EmotiW) challenge. EmotiW 2018 is a grand challenge in the ACM International Conference on Multimodal Interaction 2018, Colorado, USA. The challenge aims at p…

Emotion Recognitionmultimodal interaction