Pre-training strategies and datasets for facial representation learning
What is the best way to learn a universal face representation? Recent work on Deep Learning in the area of face analysis has focused on supervised learning for specific tasks of interest (e.g. face recognition, facial landmark localization etc.) but has overlooked the overarching question of how to find a facial representation that can be readily adapted to several facial analysis tasks and datasets. To this end, we make the following 4 contributions: (a) we introduce, for the first time, a comprehensive evaluation benchmark for facial representation learning consisting of 5 important face analysis tasks. (b) We systematically investigate two ways of large-scale representation learning applied to faces: supervised and unsupervised pre-training. Importantly, we focus our evaluations on the case of few-shot facial learning. (c) We investigate important properties of the training datasets including their size and quality (labelled, unlabelled or even uncurated). (d) To draw our conclusions, we conducted a very large number of experiments. Our main two findings are: (1) Unsupervised pre-training on completely in-the-wild, uncurated data provides consistent and, in some cases, significant accuracy improvements for all facial tasks considered. (2) Many existing facial video datasets seem to have a large amount of redundancy. We will release code, and pre-trained models to facilitate future research.
Code (2)
Tasks
3D Face Reconstruction3D Facial Landmark LocalizationArousal EstimationEmotion RecognitionFace AlignmentFace RecognitionFacial Action Unit DetectionFacial Expression Recognition (FER)Few-Shot LearningRepresentation LearningUnsupervised Pre-trainingValence EstimationValene EstimationSimilar Papers 제목 키워드 기반
Improving Makeup Face Verification by Exploring Part-Based Representations
Recently, we have seen an increase in the global facial recognition market size. Despite significant advances in face recognition technology with the adoption of convolutional neural networks, there are still open challe…
Face RecognitionFace VerificationRevisiting Self-Supervised Contrastive Learning for Facial Expression Recognition
The success of most advanced facial expression recognition works relies heavily on large-scale annotated datasets. However, it poses great challenges in acquiring clean and consistent annotations for facial expression da…
Contrastive LearningFacial Expression RecognitionFacial Expression Recognition (FER)Self-Supervised LearningEvaluation of Self-taught Learning-based Representations for Facial Emotion Recognition
This work describes different strategies to generate unsupervised representations obtained through the concept of self-taught learning for facial emotion recognition (FER). The idea is to create complementary representat…
DiversityEmotion RecognitionFacial Emotion RecognitionFSFM: A Generalizable Face Security Foundation Model via Self-Supervised Facial Representation Learning
This work asks: with abundant, unlabeled real faces, how to learn a robust and transferable facial representation that boosts various face security tasks with respect to generalization performance? We make the first atte…
DeepFake Detectiondiffusion-generated faces detectionFace Anti-SpoofingFace Swapping+4Memory Integrity of CNNs for Cross-Dataset Facial Expression Recognition
Facial expression recognition is a major problem in the domain of artificial intelligence. One of the best ways to solve this problem is the use of convolutional neural networks (CNNs). However, a large amount of data is…
Facial Expression RecognitionFacial Expression Recognition (FER)