Learning and Recognizing Human Action from Skeleton Movement with Deep Residual Neural Networks
Automatic human action recognition is indispensable for almost artificial intelligent systems such as video surveillance, human-computer interfaces, video retrieval, etc. Despite a lot of progress, recognizing actions in an unknown video is still a challenging task in computer vision. Recently, deep learning algorithms have proved its great potential in many vision-related recognition tasks. In this paper, we propose the use of Deep Residual Neural Networks (ResNets) to learn and recognize human action from skeleton data provided by Kinect sensor. Firstly, the body joint coordinates are transformed into 3D-arrays and saved in RGB images space. Five different deep learning models based on ResNet have been designed to extract image features and classify them into classes. Experiments are conducted on two public video datasets for human action recognition containing various challenges. The results show that our method achieves the state-of-the-art performance comparing with existing approaches.
Code (0)
등록된 구현이 없습니다.
Tasks
Action RecognitionDeep LearningRetrievalTemporal Action LocalizationVideo RetrievalMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Fine-Grained Side Information Guided Dual-Prompts for Zero-Shot Skeleton Action Recognition
Skeleton-based zero-shot action recognition aims to recognize unknown human actions based on the learned priors of the known skeleton-based actions and a semantic descriptor space shared by both known and unknown categor…
Action RecognitionAttributeZero-Shot Action RecognitionZero Shot Skeletal Action RecognitionSkeletal Movement to Color Map: A Novel Representation for 3D Action Recognition with Inception Residual Networks
We propose a novel skeleton-based representation for 3D action recognition in videos using Deep Convolutional Neural Networks (D-CNNs). Two key issues have been addressed: First, how to construct a robust representation …
3D Action RecognitionAction RecognitionAction Recognition In VideosTemporal Action LocalizationKinect Sensor Based Gesture Recognition for Surveillance Application
Hand gesture recognition has been granted as one of the emerging fields in research today providing a natural way of communication between man and a machine. Gestures are some forms of body motions which a person express…
Gesture RecognitionHand Gesture RecognitionHand-Gesture RecognitionLearning to Recognize 3D Human Action from A New Skeleton-based Representation Using Deep Convolutional Neural Networks
Recognizing human actions in untrimmed videos is an important challenging task. An effective 3D motion representation and a powerful learning model are two key factors influencing recognition performance. In this paper w…
3D Action RecognitionAction RecognitionAction Recognition In VideosTemporal Action LocalizationRecognizing American Sign Language Manual Signs from RGB-D Videos
In this paper, we propose a 3D Convolutional Neural Network (3DCNN) based multi-stream framework to recognize American Sign Language (ASL) manual signs (consisting of movements of the hands, as well as non-manual face mo…