paper-with-me

홈 › Papers

Multi-modal Multi-label Facial Action Unit Detection with Transformer

2022-03-24 · Lingfeng Wang, Shisen Wang, Jin Qi

Facial Action Coding System is an important approach of facial expression analysis.This paper describes our submission to the third Affective Behavior Analysis (ABAW) 2022 competition. We proposed a transfomer based model to detect facial action unit (FAU) in video. To be specific, we firstly trained a multi-modal model to extract both audio and visual feature. After that, we proposed a action units correlation module to learn relationships between each action unit labels and refine action unit detection result. Experimental results on validation dataset shows that our method achieves better performance than baseline model, which verifies that the effectiveness of proposed network.

📄 PDF Abstract BibTeX arXiv:2203.13301

Code (0)

등록된 구현이 없습니다.

Tasks

Action Unit DetectionFacial Action Unit Detection

Similar Papers 제목 키워드 기반

MAUGen: A Unified Diffusion Approach for Multi-Identity Facial Expression and AU Label Generation

2026-01-31 · Xiangdong Li, Ye Lou, Ao Gao, Wei Zhang 외 arxiv

The lack of large-scale, demographically diverse face images with precise Action Unit (AU) occurrence and intensity annotations has long been recognized as a fundamental bottleneck in developing generalizable AU recognit…

Representation Learning

Multi-Task Multi-Modal Self-Supervised Learning for Facial Expression Recognition

2024-04-16 · Marah Halawa, Florian Blume, Pia Bideau, Martin Maier 외

Human communication is multi-modal; e.g., face-to-face interaction involves auditory signals (speech) and visual signals (face movements and hand gestures). Hence, it is essential to exploit multiple modalities when desi…

Emotion ClassificationEmotion Recognition in ConversationFacial Expression RecognitionSelf-Supervised Learning

REACT2023: the first Multi-modal Multiple Appropriate Facial Reaction Generation Challenge

2023-06-11 · Siyang Song, Micol Spitale, Cheng Luo, German Barquero 외

The Multi-modal Multiple Appropriate Facial Reaction Generation Challenge (REACT2023) is the first competition event focused on evaluating multimedia processing and machine learning techniques for generating human-approp…

Media2Face: Co-speech Facial Animation Generation With Multi-Modality Guidance

2024-01-28 · Qingcheng Zhao, Pengyu Long, Qixuan Zhang, Dafei Qin 외

The synthesis of 3D facial animations from speech has garnered considerable attention. Due to the scarcity of high-quality 4D facial data and well-annotated abundant multi-modality labels, previous methods often suffer f…

Fusing Body Posture with Facial Expressions for Joint Recognition of Affect in Child-Robot Interaction

2019-01-07 · Panagiotis P. Filntisis, Niki Efthymiou, Petros Koutras, Gerasimos Potamianos 외

In this paper we address the problem of multi-cue affect recognition in challenging scenarios such as child-robot interaction. Towards this goal we propose a method for automatic recognition of affect that leverages body…