paper-with-me

Papers

Enhancing Face Recognition With Self-Supervised 3D Reconstruction

2022-01-01 · CVPR 2022 1 · Mingjie He, Jie Zhang, Shiguang Shan, Xilin Chen

Attributed to both the development of deep networks and abundant data, automatic face recognition (FR) has quickly reached human-level capacity in the past few years. However, the FR problem is not perfectly solved in case of uncontrolled illumination and pose. In this paper, we propose to enhance face recognition with a bypass of self-supervised 3D reconstruction, which enforces the neural backbone to focus on the identity-related depth and albedo information while neglects the identity-irrelevant pose and illumination information. Specifically, inspired by the physical model of image formation, we improve the backbone FR network by introducing a 3D face reconstruction loss with two auxiliary networks. The first one estimates the pose and illumination from the input face image while the second one decodes the canonical depth and albedo from the intermediate feature of the FR backbone network. The whole network is trained in end-to-end manner with both classic face identification loss and the loss of 3D face reconstruction with the physical parameters. In this way, the self-supervised reconstruction acts as a regularization that enables the recognition network to understand faces in 3D view, and the learnt features are forced to encode more information of canonical facial depth and albedo, which is more intrinsic and beneficial to face recognition. Extensive experimental results on various face recognition benchmarks show that, without any cost of extra annotations and computations, our method outperforms state-of-the-art ones. Moreover, the learnt representations can also well generalize to other face-related downstream tasks such as the facial attribute recognition with limited labeled data.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

3D Face Reconstruction3D ReconstructionAttributeFace IdentificationFace RecognitionFace Reconstruction

Similar Papers 제목 키워드 기반

EEG-ReMinD: Enhancing Neurodegenerative EEG Decoding through Self-Supervised State Reconstruction-Primed Riemannian Dynamics

2025-01-14 · ZiRui Wang, Zhenxi Song, Yi Guo, Yuxin Liu 외

The development of EEG decoding algorithms confronts challenges such as data sparsity, subject variability, and the need for precise annotations, all of which are vital for advancing brain-computer interfaces and enhanci…

EEGEeg Decoding

Multi-Task Multi-Modal Self-Supervised Learning for Facial Expression Recognition

2024-04-16 · Marah Halawa, Florian Blume, Pia Bideau, Martin Maier 외

Human communication is multi-modal; e.g., face-to-face interaction involves auditory signals (speech) and visual signals (face movements and hand gestures). Hence, it is essential to exploit multiple modalities when desi…

Emotion ClassificationEmotion Recognition in ConversationFacial Expression RecognitionSelf-Supervised Learning

Does Visual Self-Supervision Improve Learning of Speech Representations for Emotion Recognition?

2020-05-04 · Abhinav Shukla, Stavros Petridis, Maja Pantic

Self-supervised learning has attracted plenty of recent research interest. However, most works for self-supervision in speech are typically unimodal and there has been limited work that studies the interaction between au…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Emotion RecognitionFace Reconstruction+5

Towards Metrical Reconstruction of Human Faces

2022-04-13 · Wojciech Zielonka, Timo Bolkart, Justus Thies

Face reconstruction and tracking is a building block of numerous applications in AR/VR, human-machine interaction, as well as medical applications. Most of these applications rely on a metrically correct prediction of th…

2k3D Face ReconstructionFace RecognitionFace Reconstruction

Enhancing Self-Supervised Fine-Grained Video Object Tracking with Dynamic Memory Prediction

2025-04-30 · Zihan Zhou, Changrui Dai, Aibo Song, Xiaolin Fang

Successful video analysis relies on accurate recognition of pixels across frames, and frame reconstruction methods based on video correspondence learning are popular due to their efficiency. Existing frame reconstruction…

Decision MakingObjectObject TrackingSemantic Segmentation+1