paper-with-me

홈 › Papers

Investigation of Multimodal Features, Classifiers and Fusion Methods for Emotion Recognition

2018-09-13 · Zheng Lian, Ya Li, Jian-Hua Tao, Jian Huang

Automatic emotion recognition is a challenging task. In this paper, we present our effort for the audio-video based sub-challenge of the Emotion Recognition in the Wild (EmotiW) 2018 challenge, which requires participants to assign a single emotion label to the video clip from the six universal emotions (Anger, Disgust, Fear, Happiness, Sad and Surprise) and Neutral. The proposed multimodal emotion recognition system takes audio, video and text information into account. Except for handcraft features, we also extract bottleneck features from deep neutral networks (DNNs) via transfer learning. Both temporal classifiers and non-temporal classifiers are evaluated to obtain the best unimodal emotion classification result. Then possibilities are extracted and passed into the Beam Search Fusion (BS-Fusion). We test our method in the EmotiW 2018 challenge and we gain promising results. Compared with the baseline system, there is a significant improvement. We achieve 60.34% accuracy on the testing dataset, which is only 1.5% lower than the winner. It shows that our method is very competitive.

📄 PDF Abstract BibTeX arXiv:1809.06225

Code (1)

zeroQiaoba/EmotiW2018 공식 구현 tf

Tasks

Emotion ClassificationEmotion RecognitionMultimodal Emotion RecognitionTransfer Learning

Similar Papers 제목 키워드 기반

Multimodal Industrial Anomaly Detection via Hybrid Fusion

2023-03-01 · CVPR 2023 1 · Yue Wang, Jinlong Peng, Jiangning Zhang, Ran Yi 외

2D-based Industrial Anomaly Detection has been widely discussed, however, multimodal industrial anomaly detection based on 3D point clouds and RGB images still has many untouched fields. Existing multimodal industrial an…

3D Anomaly DetectionAnomaly DetectionContrastive LearningRGB+3D Anomaly Detection and Segmentation

LLaDA-V: Large Language Diffusion Models with Visual Instruction Tuning

2025-05-22 · Zebin You, Shen Nie, Xiaolu Zhang, Jun Hu 외

In this work, we introduce LLaDA-V, a purely diffusion-based Multimodal Large Language Model (MLLM) that integrates visual instruction tuning with masked diffusion models, representing a departure from the autoregressive…

Language ModelingLanguage ModellingLarge Language ModelMultimodal Large Language Model

AIMDiT: Modality Augmentation and Interaction via Multimodal Dimension Transformation for Emotion Recognition in Conversations

2024-04-12 · Sheng Wu, Jiaxing Liu, Longbiao Wang, Dongxiao He 외

Emotion Recognition in Conversations (ERC) is a popular task in natural language processing, which aims to recognize the emotional state of the speaker in conversations. While current research primarily emphasizes contex…

Emotion RecognitionRepresentation Learning

Learning and Fusing Multimodal Features from and for Multi-task Facial Computing

2016-10-14 · Wei Li, Zhigang Zhu

We propose a deep learning-based feature fusion approach for facial computing including face recognition as well as gender, race and age detection. Instead of training a single classifier on face images to classify them …

Face Recognition

Multimodal Ensemble with Conditional Feature Fusion for Dysgraphia Diagnosis in Children from Handwriting Samples

2024-08-25 · Jayakanth Kunhoth, Somaya Al-Maadeed, Moutaz Saleh, Younes Akbari

Developmental dysgraphia is a neurological disorder that hinders children's writing skills. In recent years, researchers have increasingly explored machine learning methods to support the diagnosis of dysgraphia based on…

Diagnostic