Attention-Aware Deep Reinforcement Learning for Video Face Recognition
In this paper, we propose an attention-aware deep reinforcement learning (ADRL) method for video face recognition, which aims to discard the misleading and confounding frames and find the focuses of attention in face videos for person recognition. We formulate the process of finding the attentions of videos as a Markov decision process and train the attention model through a deep reinforcement learning framework without using extra labels. Unlike existing attention models, our method takes information from both the image space and the feature space as the input to make better use of face information that is discarded in the feature learning process. Besides, our approach is attention-aware, which seeks different attentions of videos for the verification of different pairs of videos. Our approach achieves very competitive video face recognition performance on three widely used video face datasets.
Code (0)
등록된 구현이 없습니다.
Tasks
Deep Reinforcement LearningFace RecognitionPerson Recognitionreinforcement-learningReinforcement LearningReinforcement Learning (RL)Similar Papers 제목 키워드 기반
Dependency-aware Attention Control for Unconstrained Face Recognition with Image Sets
This paper targets the problem of image set-based face verification and identification. Unlike traditional single media (an image or video) setting, we encounter a set of heterogeneous contents containing orderless image…
Face RecognitionFace VerificationReinforcement LearningAttention-Set based Metric Learning for Video Face Recognition
Face recognition has made great progress with the development of deep learning. However, video face recognition (VFR) is still an ongoing task due to various illumination, low-resolution, pose variations and motion blur.…
Face RecognitionMetric LearningAttention-Aware Transformer-Based Aggregation Network for Video Periocular Recognition
Video periocular recognition is the task of recognizing an individual's identity based on the region around an individual's eyes. The periocular area is one of the most discriminative regions of the human face, making it…
Attention Control with Metric Learning Alignment for Image Set-based Recognition
This paper considers the problem of image set-based face verification and identification. Unlike traditional single sample (an image or a video) setting, this situation assumes the availability of a set of heterogeneous …
Face RecognitionFace VerificationMetric LearningReinforcement LearningIDSelect: A RL-Based Cost-Aware Selection Agent for Video-based Multi-Modal Person Recognition
Video-based person recognition achieves robust identification by integrating face, body, and gait. However, current systems waste computational resources by processing all modalities with fixed heavyweight ensembles rega…
Reinforcement Learning