An Audio-Video Deep and Transfer Learning Framework for Multimodal Emotion Recognition in the wild
In this paper, we present our contribution to ABAW facial expression challenge. We report the proposed system and the official challenge results adhering to the challenge protocol. Using end-to-end deep learning and benefiting from transfer learning approaches, we reached a test set challenge performance measure of 42.10%.
Code (0)
등록된 구현이 없습니다.
Tasks
Deep LearningEmotion RecognitionMultimodal Emotion RecognitionTransfer LearningSimilar Papers 제목 키워드 기반
HunyuanVideo-Avatar: High-Fidelity Audio-Driven Human Animation for Multiple Characters
Recent years have witnessed significant progress in audio-driven human animation. However, critical challenges remain in (i) generating highly dynamic videos while preserving character consistency, (ii) achieving precise…
Human AnimationInvestigation of Multimodal Features, Classifiers and Fusion Methods for Emotion Recognition
Automatic emotion recognition is a challenging task. In this paper, we present our effort for the audio-video based sub-challenge of the Emotion Recognition in the Wild (EmotiW) 2018 challenge, which requires participant…
Emotion ClassificationEmotion RecognitionMultimodal Emotion RecognitionTransfer LearningMulti-Modal Emotion Recognition by Text, Speech and Video Using Pretrained Transformers
Due to the complex nature of human emotions and the diversity of emotion representation methods in humans, emotion recognition is a challenging field. In this research, three input modalities, namely text, audio (speech)…
DiversityEmotion RecognitionMultimodal Emotion RecognitionSelf-Supervised Learning+1Retrieval-Augmented Multimodal Depression Detection
Multimodal deep learning has shown promise in depression detection by integrating text, audio, and video signals. Recent work leverages sentiment analysis to enhance emotional understanding, yet suffers from high computa…
Multimodal Deep LearningMulti-Task LearningSentiment AnalysisTransfer LearningEnhancing Modal Fusion by Alignment and Label Matching for Multimodal Emotion Recognition
To address the limitation in multimodal emotion recognition (MER) performance arising from inter-modal information fusion, we propose a novel MER framework based on multitask learning where fusion occurs after alignment,…
Contrastive LearningEmotion RecognitionMultimodal Emotion Recognition