Enhancing Ambiguous Dynamic Facial Expression Recognition with Soft Label-based Data Augmentation
Dynamic facial expression recognition (DFER) is a task that estimates emotions from facial expression video sequences. For practical applications, accurately recognizing ambiguous facial expressions -- frequently encountered in in-the-wild data -- is essential. In this study, we propose MIDAS, a data augmentation method designed to enhance DFER performance for ambiguous facial expression data using soft labels representing probabilities of multiple emotion classes. MIDAS augments training data by convexly combining pairs of video frames and their corresponding emotion class labels. This approach extends mixup to soft-labeled video data, offering a simple yet highly effective method for handling ambiguity in DFER. To evaluate MIDAS, we conducted experiments on both the DFEW dataset and FERV39k-Plus, a newly constructed dataset that assigns soft labels to an existing DFER dataset. The results demonstrate that models trained with MIDAS-augmented data achieve superior performance compared to the state-of-the-art method trained on the original dataset.
Code (0)
등록된 구현이 없습니다.
Tasks
Data AugmentationDynamic Facial Expression RecognitionFacial Expression RecognitionMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
MIDAS: Mixing Ambiguous Data with Soft Labels for Dynamic Facial Expression Recognition
Dynamic facial expression recognition (DFER) is an important task in the field of computer vision. To apply automatic DFER in practice, it is necessary to accurately recognize ambiguous facial expressions, which often ap…
Data AugmentationDynamic Facial Expression RecognitionFacial Expression RecognitionFrom Static to Dynamic: Adapting Landmark-Aware Image Models for Facial Expression Recognition in Videos
Dynamic facial expression recognition (DFER) in the wild is still hindered by data limitations, e.g., insufficient quantity and diversity of pose, occlusion and illumination, as well as the inherent ambiguity of facial e…
Dynamic Facial Expression RecognitionFacial Expression RecognitionFacial Expression Recognition (FER)Multi Modal Facial Expression Recognition with Transformer-Based Fusion Networks and Dynamic Sampling
Facial expression recognition is an essential task for various applications, including emotion detection, mental health analysis, and human-machine interactions. In this paper, we propose a multi-modal facial expression …
Facial Expression RecognitionFacial Expression Recognition (FER)Rethinking the Learning Paradigm for Facial Expression Recognition
Due to the subjective crowdsourcing annotations and the inherent inter-class similarity of facial expressions, the real-world Facial Expression Recognition (FER) datasets usually exhibit ambiguous annotation. To simplify…
Facial Expression RecognitionFacial Expression Recognition (FER)Rethinking Occlusion in FER: A Semantic-Aware Perspective and Go Beyond
Facial expression recognition (FER) is a challenging task due to pervasive occlusion and dataset biases. Especially when facial information is partially occluded, existing FER models struggle to extract effective facial …
Facial Expression RecognitionSemantic Segmentation