Attentional Feature-Pair Relation Networks for Accurate Face Recognition
Human face recognition is one of the most important research areas in biometrics. However, the robust face recognition under a drastic change of the facial pose, expression, and illumination is a big challenging problem for its practical application. Such variations make face recognition more difficult. In this paper, we propose a novel face recognition method, called Attentional Feature-pair Relation Network (AFRN), which represents the face by the relevant pairs of local appearance block features with their attention scores. The AFRN represents the face by all possible pairs of the 9x9 local appearance block features, the importance of each pair is considered by the attention map that is obtained from the low-rank bilinear pooling, and each pair is weighted by its corresponding attention score. To increase the accuracy, we select top-K pairs of local appearance block features as relevant facial information and drop the remaining irrelevant. The weighted top-K pairs are propagated to extract the joint feature-pair relation by using bilinear attention network. In experiments, we show the effectiveness of the proposed AFRN and achieve the outstanding performance in the 1:1 face verification and 1:N face identification tasks compared to existing state-of-the-art methods on the challenging LFW, YTF, CALFW, CPLFW, CFP, AgeDB, IJB-A, IJB-B, and IJB-C datasets.
Code (0)
등록된 구현이 없습니다.
Tasks
Face IdentificationFace RecognitionFace VerificationRelationRelation NetworkRobust Face RecognitionSimilar Papers 제목 키워드 기반
AAFACE: Attribute-aware Attentional Network for Face Recognition
In this paper, we present a new multi-branch neural network that simultaneously performs soft biometric (SB) prediction as an auxiliary modality and face recognition (FR) as the main task. Our proposed network named AAFa…
AttributeFace RecognitionCross Attentional Audio-Visual Fusion for Dimensional Emotion Recognition
Multimodal analysis has recently drawn much interest in affective computing, since it can improve the overall accuracy of emotion recognition over isolated uni-modal approaches. The most effective techniques for multimod…
Emotion RecognitionMultimodal Emotion RecognitionRelation Network for Multi-label Aerial Image Classification
Multi-label classification plays a momentous role in perceiving intricate contents of an aerial image and triggers several related studies over the last years. However, most of them deploy few efforts in exploiting label…
ClassificationGeneral Classificationimage-classificationImage Classification+5Predicting human gaze using low-level saliency combined with face detection
Under natural viewing conditions, human observers shift their gaze to allocate processing resources to subsets of the visual input. Many computational models have aimed at predicting such voluntary attentional shifts. Al…
Face DetectionSyncMask: Synchronized Attentional Masking for Fashion-centric Vision-Language Pretraining
Vision-language models (VLMs) have made significant strides in cross-modal understanding through large-scale paired datasets. However, in fashion domain, datasets often exhibit a disparity between the information conveye…
Contrastive LearningImage-text matchingLanguage ModelingLanguage Modelling+2