paper-with-me

홈 › Papers

AU-Aware Vision Transformers for Biased Facial Expression Recognition

2022-11-12 · Shuyi Mao, Xinpeng Li, Qingyang Wu, Xiaojiang Peng

Studies have proven that domain bias and label bias exist in different Facial Expression Recognition (FER) datasets, making it hard to improve the performance of a specific dataset by adding other datasets. For the FER bias issue, recent researches mainly focus on the cross-domain issue with advanced domain adaption algorithms. This paper addresses another problem: how to boost FER performance by leveraging cross-domain datasets. Unlike the coarse and biased expression label, the facial Action Unit (AU) is fine-grained and objective suggested by psychological studies. Motivated by this, we resort to the AU information of different FER datasets for performance boosting and make contributions as follows. First, we experimentally show that the naive joint training of multiple FER datasets is harmful to the FER performance of individual datasets. We further introduce expression-specific mean images and AU cosine distances to measure FER dataset bias. This novel measurement shows consistent conclusions with experimental degradation of joint training. Second, we propose a simple yet conceptually-new framework, AU-aware Vision Transformer (AU-ViT). It improves the performance of individual datasets by jointly training auxiliary datasets with AU or pseudo-AU labels. We also find that the AU-ViT is robust to real-world occlusions. Moreover, for the first time, we prove that a carefully-initialized ViT achieves comparable performance to advanced deep convolutional networks. Our AU-ViT achieves state-of-the-art performance on three popular datasets, namely 91.10% on RAF-DB, 65.59% on AffectNet, and 90.15% on FERPlus. The code and models will be released soon.

📄 PDF Abstract BibTeX arXiv:2211.06609

Code (0)

등록된 구현이 없습니다.

Tasks

Domain AdaptationFacial Expression RecognitionFacial Expression Recognition (FER)

Methods 이 논문이 사용한 방법론

Multi-Head Attention 설명 없음
Attention 설명 없음
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
Position-Wise Feed-Forward Layer 설명 없음
Label Smoothing Label Smoothing is a regularization technique that introduces noise for the labels. This accounts for the fact that datasets may have mistakes in them, so maximizing the…
Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Adam 설명 없음
Absolute Position Encodings Absolute Position Encodings are a type of position embeddings for [Transformer-based models] where positional encodings are…

Similar Papers 제목 키워드 기반

GReFEL: Geometry-Aware Reliable Facial Expression Learning under Bias and Imbalanced Data Distribution

2024-10-21 · Azmine Toushik Wasi, Taki Hasan Rafi, Raima Islam, Karlo Serbetar 외

Reliable facial expression learning (FEL) involves the effective learning of distinctive facial expression characteristics for more reliable, unbiased and accurate predictions in real-life settings. However, current syst…

Facial Expression RecognitionFacial Expression Recognition (FER)

Emotion Separation and Recognition from a Facial Expression by Generating the Poker Face with Vision Transformers

2022-07-22 · Jia Li, Jiantao Nie, Dan Guo, Richang Hong 외

Representation learning and feature disentanglement have garnered significant research interest in the field of facial expression recognition (FER). The inherent ambiguity of emotion labels poses challenges for conventio…

DisentanglementFace GenerationFacial Expression RecognitionFacial Expression Recognition (FER)+1

TransFER: Learning Relation-aware Facial Expression Representations with Transformers

2021-08-25 · ICCV 2021 10 · Fanglei Xue, Qiangchang Wang, Guodong Guo

Facial expression recognition (FER) has received increasing interest in computer vision. We propose the TransFER model which can learn rich relation-aware local representations. It mainly consists of three components: Mu…

Facial Expression RecognitionFacial Expression Recognition (FER)RelationTransfer Learning

AU-Supervised Convolutional Vision Transformers for Synthetic Facial Expression Recognition

2022-07-20 · Shuyi Mao, Xinpeng Li, Junyao Chen, Xiaojiang Peng

The paper describes our proposed methodology for the six basic expression classification track of Affective Behavior Analysis in-the-wild (ABAW) Competition 2022. In Learing from Synthetic Data(LSD) task, facial expressi…

Face RecognitionFacial Expression RecognitionFacial Expression Recognition (FER)

SuperFace: Preference-Aligned Facial Expression Estimation Beyond Pseudo Supervision

2026-05-07 · Zejian Kang, Xuanyang Xu, Wentao Yang, Kai Zheng 외 arxiv

Accurate facial estimation is crucial for realistic digital human animation, and ARKit blendshape coefficients offer an interpretable representation by mapping facial motions to semantic animation controls. However, lear…