paper-with-me

홈 › Papers

Engagement Measurement Based on Facial Landmarks and Spatial-Temporal Graph Convolutional Networks

2024-03-25 · Ali Abedi, Shehroz S. Khan

Engagement in virtual learning is crucial for a variety of factors including student satisfaction, performance, and compliance with learning programs, but measuring it is a challenging task. There is therefore considerable interest in utilizing artificial intelligence and affective computing to measure engagement in natural settings as well as on a large scale. This paper introduces a novel, privacy-preserving method for engagement measurement from videos. It uses facial landmarks, which carry no personally identifiable information, extracted from videos via the MediaPipe deep learning solution. The extracted facial landmarks are fed to Spatial-Temporal Graph Convolutional Networks (ST-GCNs) to output the engagement level of the student in the video. To integrate the ordinal nature of the engagement variable into the training process, ST-GCNs undergo training in a novel ordinal learning framework based on transfer learning. Experimental results on two video student engagement measurement datasets show the superiority of the proposed method compared to previous methods with improved state-of-the-art on the EngageNet dataset with a 3.1% improvement in four-class engagement level classification accuracy and on the Online Student Engagement dataset with a 1.5% improvement in binary engagement classification accuracy. Gradient-weighted Class Activation Mapping (Grad-CAM) was applied to the developed ST-GCNs to interpret the engagement measurements obtained by the proposed method in both the spatial and temporal domains. The relatively lightweight and fast ST-GCN and its integration with the real-time MediaPipe make the proposed approach capable of being deployed on virtual learning platforms and measuring engagement in real-time.

📄 PDF Abstract BibTeX arXiv:2403.17175

Code (0)

등록된 구현이 없습니다.

Tasks

Privacy PreservingTransfer Learning

Similar Papers 제목 키워드 기반

Deepfake Detection using ImageNet models and Temporal Images of 468 Facial Landmarks

2022-08-15 · Christeen T Jose

This paper presents our results and findings on the use of temporal images for deepfake detection. We modelled temporal relations that exist in the movement of 468 facial landmarks across frames of a given video as spati…

DeepFake DetectionFace Swapping

1DFormer: a Transformer Architecture Learning 1D Landmark Representations for Facial Landmark Tracking

2023-11-01 · Shi Yin, Shijie Huan, Shangfei Wang, Jinshui Hu 외

Recently, heatmap regression methods based on 1D landmark representations have shown prominent performance on locating facial landmarks. However, previous methods ignored to make deep explorations on the good potentials …

Landmark Tracking

Towards Omni-Supervised Face Alignment for Large Scale Unlabeled Videos

2019-12-16 · Congcong Zhu, Hao liu, Zhenhua Yu, Xuehong Sun

In this paper, we propose a spatial-temporal relational reasoning networks (STRRN) approach to investigate the problem of omni-supervised face alignment in videos. Unlike existing fully supervised methods which rely on n…

Face AlignmentRelational Reasoning

A Novel Space-Time Representation on the Positive Semidefinite Con for Facial Expression Recognition

2017-07-20 · ICCV 2017 · Anis Kacem, Mohamed Daoudi, Boulbaba Ben Amor, Juan Carlos Alvarez-Paiva

In this paper, we study the problem of facial expression recognition using a novel space-time geometric representation. We describe the temporal evolution of facial landmarks as parametrized trajectories on the Riemannia…

Facial Expression RecognitionFacial Expression Recognition (FER)

A Novel Space-Time Representation on the Positive Semidefinite Cone for Facial Expression Recognition

2017-10-01 · ICCV 2017 10 · Anis Kacem, Mohamed Daoudi, Boulbaba Ben Amor, Juan Carlos Alvarez-Paiva

In this paper, we study the problem of facial expression recognition using a novel space-time geometric representation. We describe the temporal evolution of facial landmarks as parametrized trajectories on the Riemannia…

Facial Expression RecognitionFacial Expression Recognition (FER)