paper-with-me

홈 › Papers

Cross-Attention Based Influence Model for Manual and Nonmanual Sign Language Analysis

2024-09-12 · Lipisha Chaudhary, Fei Xu, Ifeoma Nwogu

Both manual (relating to the use of hands) and non-manual markers (NMM), such as facial expressions or mouthing cues, are important for providing the complete meaning of phrases in American Sign Language (ASL). Efforts have been made in advancing sign language to spoken/written language understanding, but most of these have primarily focused on manual features only. In this work, using advanced neural machine translation methods, we examine and report on the extent to which facial expressions contribute to understanding sign language phrases. We present a sign language translation architecture consisting of two-stream encoders, with one encoder handling the face and the other handling the upper body (with hands). We propose a new parallel cross-attention decoding mechanism that is useful for quantifying the influence of each input modality on the output. The two streams from the encoder are directed simultaneously to different attention stacks in the decoder. Examining the properties of the parallel cross-attention weights allows us to analyze the importance of facial markers compared to body and hand features during a translating task.

📄 PDF Abstract BibTeX arXiv:2409.08162

Code (0)

등록된 구현이 없습니다.

Tasks

DecoderMachine TranslationSign Language TranslationTranslation

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음
American 설명 없음

Similar Papers 제목 키워드 기반

A survey of Shading Techniques for Facial Deformations on Sign Language Avatars

2020-05-01 · LREC 2020 5 · Ronan Johnson, Rosalee Wolfe

Of the five phonemic parameters in sign language (handshape, location, palm orientation, movement and nonmanual expressions), the one that still poses the most challenges for effective avatar display is nonmanual signals…

Recognition of Nonmanual Markers in American Sign Language (ASL) Using Non-Parametric Adaptive 2D-3D Face Tracking

2012-05-01 · LREC 2012 5 · Dimitris Metaxas, Bo Liu, Fei Yang, Peng Yang 외

This paper addresses the problem of automatically recognizing linguistically significant nonmanual expressions in American Sign Language from video. We develop a fully automatic system that is able to track facial expres…

Sign Language Recognition

Recognizing American Sign Language Nonmanual Signal Grammar Errors in Continuous Videos

2020-05-01 · Elahe Vahdani, Longlong Jing, YingLi Tian, Matt Huenerfauth

As part of the development of an educational tool that can help students achieve fluency in American Sign Language (ASL) through independent and interactive practice with immediate feedback, this paper introduces a near …

A Novel Approach to Managing Lower Face Complexity in Signing Avatars

2022-06-01 · SLTAT (LREC) 2022 6 · John McDonald, Ronan Johnson, Rosalee Wolfe

An avatar that produces legible, easy-to-understand signing is one of the essential components to an effective automatic signed/spoken translation system. Facial nonmanual signals are essential to natural signing, but un…

Translation

Testing MediaPipe Holistic for Linguistic Analysis of Nonmanual Markers in Sign Languages

2024-03-15 · Anna Kuznetsova, Vadim Kimmelman

Advances in Deep Learning have made possible reliable landmark tracking of human bodies and faces that can be used for a variety of tasks. We test a recent Computer Vision solution, MediaPipe Holistic (MPH), to find out …

Landmark Tracking