paper-with-me

홈 › Papers

A transformer-based approach to video frame-level prediction in Affective Behaviour Analysis In-the-wild

2023-03-16 · Dang-Khanh Nguyen, Ngoc-Huynh Ho, Sudarshan Pant, Hyung-Jeong Yang

In recent years, transformer architecture has been a dominating paradigm in many applications, including affective computing. In this report, we propose our transformer-based model to handle Emotion Classification Task in the 5th Affective Behavior Analysis In-the-wild Competition. By leveraging the attentive model and the synthetic dataset, we attain a score of 0.4775 on the validation set of Aff-Wild2, the dataset provided by the organizer.

📄 PDF Abstract BibTeX arXiv:2303.09293

Code (0)

등록된 구현이 없습니다.

Tasks

Emotion Classification

Similar Papers 제목 키워드 기반

Multi-Granularity Network with Modal Attention for Dense Affective Understanding

2021-06-18 · Baoming Yan, Lin Wang, Ke Gao, Bo Gao 외

Video affective understanding, which aims to predict the evoked expressions by the video content, is desired for video creation and recommendation. In the recent EEV challenge, a dense affective understanding task is pro…

Modeling Induced Pleasure through Cognitive Appraisal Prediction via Multimodal Fusion

2026-04-26 · Nastaran Dab, Raziyeh Zall, Mohammadreza Kangavari arxiv

Multimodal affective computing analyzes user-generated social media content to predict emotional states. However, a critical gap remains in understanding how visual content shapes cognitive interpretations and elicits sp…

StimuVAR: Spatiotemporal Stimuli-aware Video Affective Reasoning with Multimodal Large Language Models

2024-08-31 · Yuxiang Guo, Faizan Siddiqui, Yang Zhao, Rama Chellappa 외

Predicting and reasoning how a video would make a human feel is crucial for developing socially intelligent systems. Although Multimodal Large Language Models (MLLMs) have shown impressive video understanding capabilitie…

Video Understanding

Video2Music: Suitable Music Generation from Videos using an Affective Multimodal Transformer model

2023-11-02 · Jaeyong Kang, Soujanya Poria, Dorien Herremans

Numerous studies in the field of music generation have demonstrated impressive performance, yet virtually no models are able to directly generate music to match accompanying videos. In this work, we develop a generative …

Music GenerationRhythm

Video-Based Frame-Level Facial Analysis of Affective Behavior on Mobile Devices Using EfficientNets

2022-06-10 · CVPR Workshop 2022 6 · Savchenko A.V.

In this paper, we consider the problem of real-time video-based facial emotion analytics, namely, facial expression recognition, prediction of valence and arousal and detection of action unit points. We propose the novel…

Action Unit DetectionArousal EstimationEmotion RecognitionFacial Expression Recognition+2