paper-with-me

홈 › Papers

Learning and Recognizing Human Action from Skeleton Movement with Deep Residual Neural Networks

2018-03-21 · Huy-Hieu Pham, Louahdi Khoudour, Alain Crouzil, Pablo Zegers, Sergio A. Velastin

Automatic human action recognition is indispensable for almost artificial intelligent systems such as video surveillance, human-computer interfaces, video retrieval, etc. Despite a lot of progress, recognizing actions in an unknown video is still a challenging task in computer vision. Recently, deep learning algorithms have proved its great potential in many vision-related recognition tasks. In this paper, we propose the use of Deep Residual Neural Networks (ResNets) to learn and recognize human action from skeleton data provided by Kinect sensor. Firstly, the body joint coordinates are transformed into 3D-arrays and saved in RGB images space. Five different deep learning models based on ResNet have been designed to extract image features and classify them into classes. Experiments are conducted on two public video datasets for human action recognition containing various challenges. The results show that our method achieves the state-of-the-art performance comparing with existing approaches.

📄 PDF Abstract BibTeX arXiv:1803.07780

Code (0)

등록된 구현이 없습니다.

Tasks

Action RecognitionDeep LearningRetrievalTemporal Action LocalizationVideo Retrieval

Methods 이 논문이 사용한 방법론

Average Pooling 설명 없음
ReLU How Do I Communicate to Expedia? How Do I Communicate to Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Live Support & Special Travel…
1x1 Convolution A 1 x 1 Convolution is a convolution with some special properties in that it can be used for dimensionality reduction,…
Batch Normalization 설명 없음
Bottleneck Residual Block A Bottleneck Residual Block is a variant of the residual block that utilises 1x1 convolutions to create a bottleneck. The…
Global Average Pooling Global Average Pooling is a pooling operation designed to replace fully connected layers in classical CNNs. The idea is to generate one feature map for each corresponding…
Residual Block Residual Blocks are skip-connection blocks that learn residual functions with reference to the layer inputs, instead of learning unreferenced functions. They were introduced…
Kaiming Initialization 설명 없음

Similar Papers 제목 키워드 기반

Fine-Grained Side Information Guided Dual-Prompts for Zero-Shot Skeleton Action Recognition

2024-04-11 · Yang Chen, Jingcai Guo, Tian He, Ling Wang

Skeleton-based zero-shot action recognition aims to recognize unknown human actions based on the learned priors of the known skeleton-based actions and a semantic descriptor space shared by both known and unknown categor…

Action RecognitionAttributeZero-Shot Action RecognitionZero Shot Skeletal Action Recognition

Skeletal Movement to Color Map: A Novel Representation for 3D Action Recognition with Inception Residual Networks

2018-07-18 · Huy Hieu Pham, Louahdi Khoudour, Alain Crouzil, Pablo Zegers 외

We propose a novel skeleton-based representation for 3D action recognition in videos using Deep Convolutional Neural Networks (D-CNNs). Two key issues have been addressed: First, how to construct a robust representation …

3D Action RecognitionAction RecognitionAction Recognition In VideosTemporal Action Localization

Kinect Sensor Based Gesture Recognition for Surveillance Application

2018-12-22 · Biswarup Ganguly, Amit Konar

Hand gesture recognition has been granted as one of the emerging fields in research today providing a natural way of communication between man and a machine. Gestures are some forms of body motions which a person express…

Gesture RecognitionHand Gesture RecognitionHand-Gesture Recognition

Learning to Recognize 3D Human Action from A New Skeleton-based Representation Using Deep Convolutional Neural Networks

2018-12-26 · Huy-Hieu Pham, Louahdi Khoudour, Alain Crouzil, Pablo Zegers 외

Recognizing human actions in untrimmed videos is an important challenging task. An effective 3D motion representation and a powerful learning model are two key factors influencing recognition performance. In this paper w…

3D Action RecognitionAction RecognitionAction Recognition In VideosTemporal Action Localization

Recognizing American Sign Language Manual Signs from RGB-D Videos

2019-06-07 · Longlong Jing, Elahe Vahdani, Matt Huenerfauth, YingLi Tian

In this paper, we propose a 3D Convolutional Neural Network (3DCNN) based multi-stream framework to recognize American Sign Language (ASL) manual signs (consisting of movements of the hands, as well as non-manual face mo…