Multiresolution Match Kernels for Gesture Video Classification
The emergence of depth imaging technologies like the Microsoft Kinect has renewed interest in computational methods for gesture classification based on videos. For several years now, researchers have used the Bag-of-Features (BoF) as a primary method for generation of feature vectors from video data for recognition of gestures. However, the BoF method is a coarse representation of the information in a video, which often leads to poor similarity measures between videos. Besides, when features extracted from different spatio-temporal locations in the video are pooled to create histogram vectors in the BoF method, there is an intrinsic loss of their original locations in space and time. In this paper, we propose a new Multiresolution Match Kernel (MMK) for video classification, which can be considered as a generalization of the BoF method. We apply this procedure to hand gesture classification based on RGB-D videos of the American Sign Language(ASL) hand gestures and our results show promise and usefulness of this new method.
Code (0)
등록된 구현이 없습니다.
Tasks
ClassificationGeneral ClassificationVideo ClassificationSimilar Papers 제목 키워드 기반
Dynamic gesture retrieval: searching videos by human pose sequence
The number of static human poses is limited, it is hard to retrieve the exact videos using one single pose as the clue. However, with a pose sequence or a dynamic gesture as the keyword, retrieving specific videos become…
RetrievalAn Evaluation of Large Pre-Trained Models for Gesture Recognition using Synthetic Videos
In this work, we explore the possibility of using synthetically generated data for video-based gesture recognition with large pre-trained models. We consider whether these models have sufficiently robust and expressive r…
ClassificationGesture Recognitionzero-shot-classificationZero-Shot LearningAudio-driven Neural Gesture Reenactment with Video Motion Graphs
Human speech is often accompanied by body gestures including arm and hand gestures. We present a method that reenacts a high-quality video with gestures matching a target speech audio. The key idea of our method is to sp…
validDeep learning-based computer vision to recognize and classify suturing gestures in robot-assisted surgery
Our previous work classified a taxonomy of suturing gestures during a vesicourethral anastomosis of robotic radical prostatectomy in association with tissue tears and patient outcomes. Herein, we train deep-learning base…
ClassificationGeneral ClassificationOptical Flow EstimationJoint Skeletal and Semantic Embedding Loss for Micro-gesture Classification
In this paper, we briefly introduce the solution of our team HFUT-VUT for the Micros-gesture Classification in the MiGA challenge at IJCAI 2023. The micro-gesture classification task aims at recognizing the action catego…
Action ClassificationClassificationGesture RecognitionMicro-gesture Recognition