paper-with-me

Papers

A Real-Time Multi-Task Learning System for Joint Detection of Face, Facial Landmark and Head Pose

2023-09-21 · Qingtian Wu, Liming Zhang

Extreme head postures pose a common challenge across a spectrum of facial analysis tasks, including face detection, facial landmark detection (FLD), and head pose estimation (HPE). These tasks are interdependent, where accurate FLD relies on robust face detection, and HPE is intricately associated with these key points. This paper focuses on the integration of these tasks, particularly when addressing the complexities posed by large-angle face poses. The primary contribution of this study is the proposal of a real-time multi-task detection system capable of simultaneously performing joint detection of faces, facial landmarks, and head poses. This system builds upon the widely adopted YOLOv8 detection framework. It extends the original object detection head by incorporating additional landmark regression head, enabling efficient localization of crucial facial landmarks. Furthermore, we conduct optimizations and enhancements on various modules within the original YOLOv8 framework. To validate the effectiveness and real-time performance of our proposed model, we conduct extensive experiments on 300W-LP and AFLW2000-3D datasets. The results obtained verify the capability of our model to tackle large-angle face pose challenges while delivering real-time performance across these interconnected tasks.

📄 PDF Abstract BibTeX arXiv:2309.11773

Code (0)

등록된 구현이 없습니다.

Tasks

Face DetectionFacial Landmark DetectionHead Pose EstimationMulti-Task LearningObject DetectionPose Estimation

Methods 이 논문이 사용한 방법론

YOLOv8 설명 없음

Similar Papers 제목 키워드 기반

Hierarchical Reactive Grasping via Task-Space Velocity Fields and Joint-Space Quadratic Programming

2025-09-01 · Yonghyeon Lee, Tzu-Yuan Lin, Alexander Alexiev, Sangbae Kim arxiv

We present a fast and reactive grasping framework that combines task-space velocity fields with joint-space Quadratic Program (QP) in a hierarchical structure. Reactive, collision-free global motion planning is particula…

Motion Planning

MultiQT: Multimodal Learning for Real-Time Question Tracking in Speech

2020-05-02 · ACL 2020 6 · Jakob D. Havtorn, Jan Latko, Joakim Edin, Lasse Borgholt 외

We address a challenging and practical task of labeling questions in speech in real time during telephone calls to emergency medical services in English, which embeds within a broader decision support system for emergenc…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

sEMG-Driven Physics-Informed Gated Recurrent Networks for Modeling Upper Limb Multi-Joint Movement Dynamics

2024-08-29 · Rajnish Kumar, Anand Gupta, Suriya Prakash Muthukrishnan, Lalan Kumar 외

Exoskeletons and rehabilitation systems offer great potential for enhancing human strength and recovery through advanced human-machine interfaces (HMIs) that adapt to movement dynamics. However, the real-time application…

A Speaker Turn-Aware Multi-Task Adversarial Network for Joint User Satisfaction Estimation and Sentiment Analysis

2024-10-12 · Kaisong Song, Yangyang Kang, Jiawei Liu, Xurui Li 외

User Satisfaction Estimation is an important task and increasingly being applied in goal-oriented dialogue systems to estimate whether the user is satisfied with the service. It is observed that whether the user's needs …

Goal-Oriented Dialogue SystemsSentiment Analysis

Opinion Recommendation using Neural Memory Model

2017-02-06 · Zhongqing Wang, Yue Zhang

We present opinion recommendation, a novel task of jointly predicting a custom review with a rating score that a certain user would give to a certain product or service, given existing reviews and rating scores to the pr…

model