paper-with-me

홈 › Papers

Video-based Sign Language Recognition without Temporal Segmentation

2018-01-30 · Jie Huang, Wengang Zhou, Qilin Zhang, Houqiang Li, Weiping Li

Millions of hearing impaired people around the world routinely use some variants of sign languages to communicate, thus the automatic translation of a sign language is meaningful and important. Currently, there are two sub-problems in Sign Language Recognition (SLR), i.e., isolated SLR that recognizes word by word and continuous SLR that translates entire sentences. Existing continuous SLR methods typically utilize isolated SLRs as building blocks, with an extra layer of preprocessing (temporal segmentation) and another layer of post-processing (sentence synthesis). Unfortunately, temporal segmentation itself is non-trivial and inevitably propagates errors into subsequent steps. Worse still, isolated SLR methods typically require strenuous labeling of each word separately in a sentence, severely limiting the amount of attainable training data. To address these challenges, we propose a novel continuous sign recognition framework, the Hierarchical Attention Network with Latent Space (LS-HAN), which eliminates the preprocessing of temporal segmentation. The proposed LS-HAN consists of three components: a two-stream Convolutional Neural Network (CNN) for video feature representation generation, a Latent Space (LS) for semantic gap bridging, and a Hierarchical Attention Network (HAN) for latent space based recognition. Experiments are carried out on two large scale datasets. Experimental results demonstrate the effectiveness of the proposed framework.

📄 PDF Abstract BibTeX arXiv:1801.10111

Code (0)

등록된 구현이 없습니다.

Tasks

SegmentationSentenceSign Language Recognition

Similar Papers 제목 키워드 기반

Pose-based Sign Language Recognition using GCN and BERT

2020-12-01 · Anirudh Tunga, Sai Vidyaranya Nuthalapati, Juan Wachs

Sign language recognition (SLR) plays a crucial role in bridging the communication gap between the hearing and vocally impaired community and the rest of the society. Word-level sign language recognition (WSLR) is the fi…

Sign Language Recognition

Continuous Sign Language Recognition through a Context-Aware Generative Adversarial Network

2021-04-01 · Sensors 2021 4 · Ilias Papastratis, Kosmas Dimitropoulos, Petros Daras

Continuous sign language recognition is a weakly supervised task dealing with the identification of continuous sign gestures from video sequences, without any prior knowledge about the temporal boundaries between consecu…

Generative Adversarial NetworkSentenceSign Language RecognitionSign Language Translation

USTM: Unified Spatial and Temporal Modeling for Continuous Sign Language Recognition

2025-12-15 · Ahmed Abul Hasanaath, Hamzah Luqman arxiv

Continuous sign language recognition (CSLR) requires precise spatio-temporal modeling to accurately recognize sequences of gestures in videos. Existing frameworks often rely on CNN-based spatial backbones combined with t…

Sign Language Recognition

Video Jigsaw: Unsupervised Learning of Spatiotemporal Context for Video Action Recognition

2018-08-22 · Unaiza Ahsan, Rishi Madhok, Irfan Essa

We propose a self-supervised learning method to jointly reason about spatial and temporal context for video recognition. Recent self-supervised approaches have used spatial context [9, 34] as well as temporal coherency […

Action RecognitionActivity RecognitionOptical Flow EstimationPosition+4

SMART: MLLM-guided Temporal Alignment for Unifying Sign Language Recognition and Spotting

2026-08-26 · Eunjee Choi, JungHoon Sung, Seongwhan Cho, Chu Xin 외 arxiv

Continuous sign language recognition (CSLR) aims to recognize gloss sequences from unsegmented sign videos under weak sequence-level supervision. However, existing methods rely on sentence-level gloss annotations, provid…

Sign Language RecognitionRepresentation Learning