Skeleton based Activity Recognition by Fusing Part-wise Spatio-temporal and Attention Driven Residues
There exist a wide range of intra class variations of the same actions and inter class similarity among the actions, at the same time, which makes the action recognition in videos very challenging. In this paper, we present a novel skeleton-based part-wise Spatiotemporal CNN RIAC Network-based 3D human action recognition framework to visualise the action dynamics in part wise manner and utilise each part for action recognition by applying weighted late fusion mechanism. Part wise skeleton based motion dynamics helps to highlight local features of the skeleton which is performed by partitioning the complete skeleton in five parts such as Head to Spine, Left Leg, Right Leg, Left Hand, Right Hand. The RIAFNet architecture is greatly inspired by the InceptionV4 architecture which unified the ResNet and Inception based Spatio-temporal feature representation concept and achieving the highest top-1 accuracy till date. To extract and learn salient features for action recognition, attention driven residues are used which enhance the performance of residual components for effective 3D skeleton-based Spatio-temporal action representation. The robustness of the proposed framework is evaluated by performing extensive experiments on three challenging datasets such as UT Kinect Action 3D, Florence 3D action Dataset, and MSR Daily Action3D datasets, which consistently demonstrate the superiority of our method
Code (0)
등록된 구현이 없습니다.
Tasks
3D Action RecognitionAction RecognitionAction Recognition In VideosActivity RecognitionTemporal Action LocalizationMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Expansion-Squeeze-Excitation Fusion Network for Elderly Activity Recognition
This work focuses on the task of elderly activity recognition, which is a challenging task due to the existence of individual actions and human-object interactions in elderly activities. Thus, we attempt to effectively a…
Action RecognitionActivity RecognitionHuman-Object Interaction DetectionGaitSTR: Gait Recognition with Sequential Two-stream Refinement
Gait recognition aims to identify a person based on their walking sequences, serving as a useful biometric modality as it can be observed from long distances without requiring cooperation from the subject. In representin…
Gait RecognitionMultiview Gait RecognitionSAFER-Activities: A Dataset for Smart Assessment of Fall Events and Routine Activities
Smart healthcare monitoring systems require precise action recognition to ensure well-being and timely intervention in critical situations such as falls, particularly for mobility-challenged individuals. Existing dataset…
Action RecognitionFusing Higher-order Features in Graph Neural Networks for Skeleton-based Action Recognition
Skeleton sequences are lightweight and compact, and thus are ideal candidates for action recognition on edge devices. Recent skeleton-based action recognition methods extract features from 3D joint coordinates as spatial…
Action RecognitionGraph Neural NetworkSkeleton Based Action RecognitionThree-Stream Convolutional Neural Network With Multi-Task and Ensemble Learning for 3D Action Recognition
In this paper, we propose a three-stream convolutional neural network (3SCNN) for action recognition from skeleton sequences, which aims to thoroughly and fully exploit the skeleton data by extracting, learning, fusing a…
3D Action RecognitionAction RecognitionEnsemble LearningSkeleton Based Action Recognition