paper-with-me

홈 › Papers

SignBERT+: Hand-model-aware Self-supervised Pre-training for Sign Language Understanding

2023-05-08 · Hezhen Hu, Weichao Zhao, Wengang Zhou, Houqiang Li

Hand gesture serves as a crucial role during the expression of sign language. Current deep learning based methods for sign language understanding (SLU) are prone to over-fitting due to insufficient sign data resource and suffer limited interpretability. In this paper, we propose the first self-supervised pre-trainable SignBERT+ framework with model-aware hand prior incorporated. In our framework, the hand pose is regarded as a visual token, which is derived from an off-the-shelf detector. Each visual token is embedded with gesture state and spatial-temporal position encoding. To take full advantage of current sign data resource, we first perform self-supervised learning to model its statistics. To this end, we design multi-level masked modeling strategies (joint, frame and clip) to mimic common failure detection cases. Jointly with these masked modeling strategies, we incorporate model-aware hand prior to better capture hierarchical context over the sequence. After the pre-training, we carefully design simple yet effective prediction heads for downstream tasks. To validate the effectiveness of our framework, we perform extensive experiments on three main SLU tasks, involving isolated and continuous sign language recognition (SLR), and sign language translation (SLT). Experimental results demonstrate the effectiveness of our method, achieving new state-of-the-art performance with a notable gain.

📄 PDF Abstract BibTeX arXiv:2305.04868

Code (0)

등록된 구현이 없습니다.

Tasks

Self-Supervised LearningSign Language RecognitionSign Language Translation

Similar Papers 제목 키워드 기반

SignBERT: Pre-Training of Hand-Model-Aware Representation for Sign Language Recognition

2021-10-11 · ICCV 2021 10 · Hezhen Hu, Weichao Zhao, Wengang Zhou, Yuechen Wang 외

Hand gesture serves as a critical role in sign language. Current deep-learning-based sign language recognition (SLR) methods may suffer insufficient interpretability and overfitting due to limited sign data sources. In t…

Self-Supervised LearningSign Language Recognition

HandMIM: Pose-Aware Self-Supervised Learning for 3D Hand Mesh Estimation

2023-07-29 · Zuyan Liu, Gaojie Lin, Congyi Wang, Min Zheng 외

With an enormous number of hand images generated over time, unleashing pose knowledge from unlabeled images for supervised hand mesh estimation is an emerging yet challenging topic. To alleviate this issue, semi-supervis…

Pose EstimationregressionRepresentation LearningSelf-Supervised Learning

Geometry-Aware Self-Training for Unsupervised Domain Adaptationon Object Point Clouds

2021-08-20 · Longkun Zou, Hui Tang, Ke Chen, Kui Jia

The point cloud representation of an object can have a large geometric variation in view of inconsistent data acquisition procedure, which thus leads to domain discrepancy due to diverse and uncontrollable shape represen…

Domain AdaptationPoint Cloud ClassificationRepresentation LearningUnsupervised Domain Adaptation

Geometry-Aware Self-Training for Unsupervised Domain Adaptation on Object Point Clouds

2021-01-01 · ICCV 2021 10 · Longkun Zou, Hui Tang, Ke Chen, Kui Jia

The point cloud representation of an object can have a large geometric variation in view of inconsistent data acquisition procedure, which thus leads to domain discrepancy due to diverse and uncontrollable shape repr…

Domain AdaptationPoint Cloud ClassificationRepresentation LearningUnsupervised Domain Adaptation

Temporal-Aware Self-Supervised Learning for 3D Hand Pose and Mesh Estimation in Videos

2020-12-06 · Liangjian Chen, Shih-Yao Lin, Yusheng Xie, Yen-Yu Lin 외

Estimating 3D hand pose directly from RGB imagesis challenging but has gained steady progress recently bytraining deep models with annotated 3D poses. Howeverannotating 3D poses is difficult and as such only a few 3Dhand…

Pose EstimationSelf-Supervised Learning