Choose Settings Carefully: Comparing Action Unit detection at Different Settings Using a Large-Scale Dataset
In this paper, we investigate the impact of some of the commonly used settings for (a) preprocessing face images, and (b) classification and training, on Action Unit (AU) detection performance and complexity. We use in our investigation a large-scale dataset, consisting of ~55K videos collected in the wild for participants watching commercial ads. The preprocessing settings include scaling the face to a fixed resolution, changing the color information (RGB to gray-scale), aligning the face, and cropping AU regions, while the classification and training settings include the kind of classifier (multi-label vs. binary) and the amount of data used for training models. To the best of our knowledge, no work had investigated the effect of those settings on AU detection. In our analysis we use CNNs as our baseline classification model.
Code (0)
등록된 구현이 없습니다.
Tasks
Action Unit DetectionClassificationSimilar Papers 제목 키워드 기반
Comparing Observation and Action Representations for Deep Reinforcement Learning in $μ$RTS
This paper presents a preliminary study comparing different observation and action space representations for Deep Reinforcement Learning (DRL) in the context of Real-time Strategy (RTS) games. Specifically, we compare tw…
Deep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)Which CNNs and Training Settings to Choose for Action Unit Detection? A Study Based on a Large-Scale Dataset
In this paper we explore the influence of some frequently used Convolutional Neural Networks (CNNs), training settings, and training set structures, on Action Unit (AU) detection. Specifically, we first compare 10 differ…
Action Unit DetectionPerformance Gains of LLMs With Humans in a World of LLMs Versus Humans
Currently, a considerable research effort is devoted to comparing LLMs to a group of human experts, where the term "expert" is often ill-defined or variable, at best, in a state of constantly updating LLM releases. Witho…
EmoBox: Multilingual Multi-corpus Speech Emotion Recognition Toolkit and Benchmark
Speech emotion recognition (SER) is an important part of human-computer interaction, receiving extensive attention from both industry and academia. However, the current research field of SER has long suffered from the fo…
Cross-corpusEmotion RecognitionSpeech Emotion Recognition``All I know about politics is what I read in Twitter'': Weakly Supervised Models for Extracting Politicians' Stances From Twitter
During the 2016 United States presidential election, politicians have increasingly used Twitter to express their beliefs, stances on current political issues, and reactions concerning national and international events. G…
All