Deep Multi-task Multi-label CNN for Effective Facial Attribute Classification
Facial Attribute Classification (FAC) has attracted increasing attention in computer vision and pattern recognition. However, state-of-the-art FAC methods perform face detection/alignment and FAC independently. The inherent dependencies between these tasks are not fully exploited. In addition, most methods predict all facial attributes using the same CNN network architecture, which ignores the different learning complexities of facial attributes. To address the above problems, we propose a novel deep multi-task multi-label CNN, termed DMM-CNN, for effective FAC. Specifically, DMM-CNN jointly optimizes two closely-related tasks (i.e., facial landmark detection and FAC) to improve the performance of FAC by taking advantage of multi-task learning. To deal with the diverse learning complexities of facial attributes, we divide the attributes into two groups: objective attributes and subjective attributes. Two different network architectures are respectively designed to extract features for two groups of attributes, and a novel dynamic weighting scheme is proposed to automatically assign the loss weight to each facial attribute during training. Furthermore, an adaptive thresholding strategy is developed to effectively alleviate the problem of class imbalance for multi-label learning. Experimental results on the challenging CelebA and LFWA datasets show the superiority of the proposed DMM-CNN method compared with several state-of-the-art FAC methods.
Code (0)
등록된 구현이 없습니다.
Tasks
AttributeFace DetectionFacial Attribute ClassificationFacial Landmark DetectionGeneral ClassificationMulti-Label LearningMulti-Task LearningSimilar Papers 제목 키워드 기반
Multi-label Learning Based Deep Transfer Neural Network for Facial Attribute Classification
Deep Neural Network (DNN) has recently achieved outstanding performance in a variety of computer vision tasks, including facial attribute classification. The great success of classifying facial attributes with DNN often …
AttributeClassificationDomain AdaptationFace Detection+5Facial Emotion Recognition with Noisy Multi-task Annotations
Human emotions can be inferred from facial expressions. However, the annotations of facial expressions are often highly noisy in common emotion coding models, including categorical and dimensional ones. To reduce human l…
Emotion RecognitionFacial Emotion RecognitionMOON: A Mixed Objective Optimization Network for the Recognition of Facial Attributes
Attribute recognition, particularly facial, extracts many labels for each image. While some multi-task vision problems can be decomposed into separate tasks and stages, e.g., training independent models for each task, fo…
AttributeAttribute ExtractionFace RecognitionMIDAS: Mixing Ambiguous Data with Soft Labels for Dynamic Facial Expression Recognition
Dynamic facial expression recognition (DFER) is an important task in the field of computer vision. To apply automatic DFER in practice, it is necessary to accurately recognize ambiguous facial expressions, which often ap…
Data AugmentationDynamic Facial Expression RecognitionFacial Expression RecognitionEnhancing Ambiguous Dynamic Facial Expression Recognition with Soft Label-based Data Augmentation
Dynamic facial expression recognition (DFER) is a task that estimates emotions from facial expression video sequences. For practical applications, accurately recognizing ambiguous facial expressions -- frequently encount…
Data AugmentationDynamic Facial Expression RecognitionFacial Expression Recognition