CLERA: A Unified Model for Joint Cognitive Load and Eye Region Analysis in the Wild
Non-intrusive, real-time analysis of the dynamics of the eye region allows us to monitor humans' visual attention allocation and estimate their mental state during the performance of real-world tasks, which can potentially benefit a wide range of human-computer interaction (HCI) applications. While commercial eye-tracking devices have been frequently employed, the difficulty of customizing these devices places unnecessary constraints on the exploration of more efficient, end-to-end models of eye dynamics. In this work, we propose CLERA, a unified model for Cognitive Load and Eye Region Analysis, which achieves precise keypoint detection and spatiotemporal tracking in a joint-learning framework. Our method demonstrates significant efficiency and outperforms prior work on tasks including cognitive load estimation, eye landmark detection, and blink estimation. We also introduce a large-scale dataset of 30k human faces with joint pupil, eye-openness, and landmark annotation, which aims to support future HCI research on human factors and eye-related analysis.
Code (0)
등록된 구현이 없습니다.
Tasks
Blink estimationKeypoint DetectionSimilar Papers 제목 키워드 기반
Eye Sclera for Fair Face Image Quality Assessment
Fair operational systems are crucial in gaining and maintaining society's trust in face recognition systems (FRS). FRS start with capturing an image and assessing its quality before using it further for enrollment or ver…
Face Image QualityFace Image Quality AssessmentFace RecognitionImage Quality AssessmentOpenEDS: Open Eye Dataset
We present a large scale data set, OpenEDS: Open Eye Dataset, of eye-images captured using a virtual-reality (VR) head mounted display mounted with two synchronized eyefacing cameras at a frame rate of 200 Hz under contr…
Semantic SegmentationClaw U-Net: A Unet-based Network with Deep Feature Concatenation for Scleral Blood Vessel Segmentation
Sturge-Weber syndrome (SWS) is a vascular malformation disease, and it may cause blindness if the patient's condition is severe. Clinical results show that SWS can be divided into two types based on the characteristics o…
SIP-SegNet: A Deep Convolutional Encoder-Decoder Network for Joint Semantic Segmentation and Extraction of Sclera, Iris and Pupil based on Periocular Region Suppression
The current developments in the field of machine vision have opened new vistas towards deploying multimodal biometric recognition systems in various real-world applications. These systems have the ability to deal with th…
DecoderDenoisingImage EnhancementReflection Removal+2Joint Iris Segmentation and Localization Using Deep Multi-task Learning Framework
Iris segmentation and localization in non-cooperative environment is challenging due to illumination variations, long distances, moving subjects and limited user cooperation, etc. Traditional methods often suffer from po…
DecoderIris SegmentationMedical Image SegmentationMulti-Task Learning+1