Boosting of Head Pose Estimation by Knowledge Distillation
We propose a response-based method of knowledge distillation (KD) for the head pose estimation problem. A student model trained by the proposed KD achieves results better than a teacher model, which is atypical for the response-based method. Our method consists of two stages. In the first stage, we trained the base neural network (NN), which has one regression head and four regression via classification (RvC) heads. We build the convolutional ensemble over the base NN using offsets of face bounding boxes over a regular grid. In the second stage, we perform KD from the convolutional ensemble into the final NN with one RvC head. The KD improves the results by an average of 7.7\% compared to base NN. This feature makes it possible to use KD as a booster and effectively train deeper NNs. NNs trained by our KD method partially improved the state-of-the-art results. KD-ResNet152 has the best results, and KD-ResNet18 has a better result on the AFLW2000 dataset than any previous method.We have made publicly available trained NNs and face bounding boxes for the 300W-LP, AFLW, AFLW2000, and BIWI datasets.Our method potentially can be effective for other regression problems.
Code (0)
등록된 구현이 없습니다.
Tasks
Head Pose EstimationKnowledge DistillationPose EstimationregressionMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
ADU-Depth: Attention-based Distillation with Uncertainty Modeling for Depth Estimation
Monocular depth estimation is challenging due to its inherent ambiguity and ill-posed nature, yet it is quite important to many applications. While recent works achieve limited accuracy by designing increasingly complica…
3D geometryDepth EstimationDomain AdaptationKnowledge Distillation+2Attention is all you need for boosting graph convolutional neural network
Graph Convolutional Neural Networks (GCNs) possess strong capabilities for processing graph data in non-grid domains. They can capture the topological logical structure and node features in graphs and integrate them into…
AllKnowledge DistillationRecommendation SystemsFast-HaMeR: Boosting Hand Mesh Reconstruction using Knowledge Distillation
Fast and accurate 3D hand reconstruction is essential for real-time applications in VR/AR, human-computer interaction, robotics, and healthcare. Most state-of-the-art methods rely on heavy models, limiting their use on r…
Knowledge DistillationBoostingBERT:Integrating Multi-Class Boosting into BERT for NLP Tasks
As a pre-trained Transformer model, BERT (Bidirectional Encoder Representations from Transformers) has achieved ground-breaking performance on multiple NLP tasks. On the other hand, Boosting is a popular ensemble learnin…
Ensemble LearningKnowledge DistillationCross-Domain Knowledge Distillation for Low-Resolution Human Pose Estimation
In practical applications of human pose estimation, low-resolution inputs frequently occur, and existing state-of-the-art models perform poorly with low-resolution images. This work focuses on boosting the performance of…
Knowledge DistillationPose Estimation