paper-with-me

Papers Video Emotion Recognition

“Video Emotion Recognition” 태그가 달린 논문 31편 · 필터 해제

CLIP-AUTT: Test-Time Personalization with Action Unit Prompting for Fine-Grained Video Emotion Recognition

2026-03-30 · Muhammad Osama Zeeshan, Masoumeh Sharafi, Benoit Savary, Alessandro Lameiras Koerich 외 arxiv

Personalization in emotion recognition (ER) is essential for accurate interpretation of subtle and subject-specific expressive patterns. Recent advances in vision-language models (VLMs), such as CLIP, demonstrate strong …

Facial Expression RecognitionVideo Emotion Recognition

Multimodal Video Emotion Recognition with Reliable Reasoning Priors

2025-07-29 · Zhepeng Wang, Yingjian Zhu, Guanghao Dong, Hongzhu Yi 외 arxiv

This study investigates the integration of trustworthy prior reasoning knowledge from MLLMs into multimodal emotion recognition. We employ Gemini to generate fine-grained, modality-separable reasoning traces, which are i…

Multimodal Emotion RecognitionVideo Emotion RecognitionContrastive Learning

VAEmo: Efficient Representation Learning for Visual-Audio Emotion with Knowledge Injection

2025-05-05 · Hao Cheng, Zhiwei Zhao, Yichao He, Zhenzhen Hu 외

Audiovisual emotion recognition (AVER) aims to infer human emotions from nonverbal visual-audio (VA) cues, offering modality-complementary and language-agnostic advantages. However, AVER remains challenging due to the in…

Contrastive LearningDynamic Facial Expression RecognitionEmotion RecognitionRepresentation Learning+2

Visual and textual prompts for enhancing emotion recognition in video

2025-04-24 · Zhifeng Wang, Qixuan Zhang, Peter Zhang, Wenjia Niu 외

Vision Large Language Models (VLLMs) exhibit promising potential for multi-modal understanding, yet their application to video-based emotion recognition remains limited by insufficient spatial and contextual awareness. T…

Emotion RecognitionVideo Emotion RecognitionVisual Prompting

MTCAE-DFER: Multi-Task Cascaded Autoencoder for Dynamic Facial Expression Recognition

2024-12-25 · Peihao Xiang, Kaida Wu, Chaohao Lin, Ou Bai

This paper expands the cascaded network branch of the autoencoder-based multi-task learning (MTL) framework for dynamic facial expression recognition, namely Multi-Task Cascaded Autoencoder for Dynamic Facial Expression …

Dynamic Facial Expression RecognitionFace DetectionFacial Emotion RecognitionFacial Expression Recognition+2

VEMOCLAP: A video emotion classification web application

2024-10-22 · Serkan Sulun, Paula Viana, Matthew E. P. Davies

We introduce VEMOCLAP: Video EMOtion Classifier using Pretrained features, the first readily available and open-source web application that analyzes the emotional content of any user-provided video. We improve our previo…

ClassificationEmotion ClassificationVideo Emotion Recognition

FE-Adapter: Adapting Image-based Emotion Classifiers to Videos

2024-08-05 · Shreyank N Gowda, Boyan Gao, David A. Clifton

Utilizing large pre-trained models for specific tasks has yielded impressive results. However, fully fine-tuning these increasingly large models is becoming prohibitively resource-intensive. This has led to a focus on mo…

Dynamic Facial Expression RecognitionEmotion RecognitionTransfer LearningVideo Emotion Recognition+1

MultiMAE-DER: Multimodal Masked Autoencoder for Dynamic Emotion Recognition

2024-04-28 · Peihao Xiang, Chaohao Lin, Kaida Wu, Ou Bai

This paper presents a novel approach to processing multimodal data for dynamic emotion recognition, named as the Multimodal Masked Autoencoder for Dynamic Emotion Recognition (MultiMAE-DER). The MultiMAE-DER leverages th…

Emotion RecognitionMultimodal Emotion RecognitionSelf-Supervised LearningVideo Emotion Recognition

MART: Masked Affective RepresenTation Learning via Masked Temporal Distribution Distillation

2024-01-01 · CVPR 2024 1 · Zhicheng Zhang, Pancheng Zhao, Eunil Park, Jufeng Yang

Limited training data is a long-standing problem for video emotion analysis (VEA). Existing works leverage the power of large-scale image datasets for transferring while failing to extract the temporal correlation of…

Emotion RecognitionMultimodal Emotion RecognitionMultimodal Sentiment AnalysisRepresentation Learning+2

Towards Emotion Analysis in Short-form Videos: A Large-Scale Dataset and Baseline

2023-11-29 · Xuecheng Wu, Heli Sun, Junxiao Xue, Jiayu Nie 외

Nowadays, short-form videos (SVs) are essential to web information acquisition and sharing in our daily life. The prevailing use of SVs to spread emotions leads to the necessity of conducting video emotion analysis (VEA)…

audio-visual learningFormMultimodal Emotion RecognitionVideo Emotion Recognition

Affective Video Content Analysis: Decade Review and New Perspectives

2023-10-26 · Junxiao Xue, Jie Wang, Xuecheng Wu, Qian Zhang

Video content is rich in semantics and has the ability to evoke various emotions in viewers. In recent years, with the rapid development of affective computing and the explosive growth of visual data, affective video con…

Emotional IntelligenceEmotion RecognitionFacial Expression RecognitionVideo Emotion Recognition

Fuzzy Approach for Audio-Video Emotion Recognition in Computer Games for Children

2023-08-31 · Pavel Kozlov, Alisher Akram, Pakizar Shamoi

Computer games are widespread nowadays and enjoyed by people of all ages. But when it comes to kids, playing these games can be more than just fun, it is a way for them to develop important skills and build emotional int…

Audio Emotion RecognitionEmotional IntelligenceEmotion RecognitionVideo Emotion Recognition

Versatile audio-visual learning for emotion recognition

2023-05-12 · Lucas Goncalves, Seong-Gyun Leem, Wei-Cheng Lin, Berrak Sisman 외

Most current audio-visual emotion recognition models lack the flexibility needed for deployment in practical applications. We envision a multimodal system that works even when only one modality is available and can be im…

Arousal EstimationAttributeaudio-visual learningEmotion Classification+6

Weakly Supervised Video Emotion Detection and Prediction via Cross-Modal Temporal Erasing Network

2023-01-01 · CVPR 2023 1 · Zhicheng Zhang, Lijuan Wang, Jufeng Yang

Automatically predicting the emotions of user-generated videos (UGVs) receives increasing interest recently. However, existing methods mainly focus on a few key visual frames, which may limit their capacity to encode…

Video Emotion DetectionVideo Emotion Recognition

Representation Learning through Multimodal Attention and Time-Sync Comments for Affective Video Content Analysis

2022-10-14 · ACM MM22 2022 10 · Jicai Pan, Shangfei Wang, Lin Fang

Although temporal patterns inherent in visual and audio signals are crucial for affective video content analysis, they have not been thoroughly explored yet. In this paper, we propose a novel Temporal-Aware Multimodal (T…

Representation LearningVideo Emotion Recognition

FV2ES: A Fully End2End Multimodal System for Fast Yet Effective Video Emotion Recognition Inference

2022-09-21 · Qinglan Wei, Xuling Huang, Yuan Zhang

In the latest social networks, more and more people prefer to express their emotions in videos through text, speech, and rich facial expressions. Multimodal video emotion analysis techniques can help understand users' in…

Emotion RecognitionMultimodal Emotion RecognitionVideo Emotion Recognition

ICANet: A Method of Short Video Emotion Recognition Driven by Multimodal Data

2022-08-24 · Xuecheng Wu, Mengmeng Tian, Lanhang Zhai

With the fast development of artificial intelligence and short videos, emotion recognition in short videos has become one of the most important research topics in human-computer interaction. At present, most emotion reco…

Emotion RecognitionOptical Flow EstimationVideo Emotion Recognition

Multi-modal Residual Perceptron Network for Audio-Video Emotion Recognition

2021-07-21 · Xin Chang, Władysław Skarbek

Audio-Video Emotion Recognition is now attacked with Deep Neural Network modeling tools. In published papers, as a rule, the authors show only cases of the superiority in multi-modality over audio-only or video-only moda…

Emotion RecognitionVideo Emotion Recognition

Technical Report for Valence-Arousal Estimation on Affwild2 Dataset

2021-05-04 · I-Hsuan Li

In this work, we describe our method for tackling the valence-arousal estimation challenge from ABAW FG-2020 Competition. The competition organizers provide an in-the-wild Aff-Wild2 dataset for participants to analyze af…

Arousal EstimationEmotion RecognitionVideo Emotion Recognition

Exploring Emotion Features and Fusion Strategies for Audio-Video Emotion Recognition

2020-12-27 · Hengshun Zhou, Debin Meng, Yuanyuan Zhang, Xiaojiang Peng 외

The audio-video based emotion recognition aims to classify a given video into basic emotions. In this paper, we describe our approaches in EmotiW 2019, which mainly explores emotion features and feature fusion strategies…

Emotion RecognitionFacial Expression Recognition (FER)Video Emotion Recognition
1–20 / 31 다음 →