paper-with-me

홈 › Papers

Skeleton based Activity Recognition by Fusing Part-wise Spatio-temporal and Attention Driven Residues

2019-12-02 · Chhavi Dhiman, Dinesh Kumar Vishwakarma, Paras Aggarwal

There exist a wide range of intra class variations of the same actions and inter class similarity among the actions, at the same time, which makes the action recognition in videos very challenging. In this paper, we present a novel skeleton-based part-wise Spatiotemporal CNN RIAC Network-based 3D human action recognition framework to visualise the action dynamics in part wise manner and utilise each part for action recognition by applying weighted late fusion mechanism. Part wise skeleton based motion dynamics helps to highlight local features of the skeleton which is performed by partitioning the complete skeleton in five parts such as Head to Spine, Left Leg, Right Leg, Left Hand, Right Hand. The RIAFNet architecture is greatly inspired by the InceptionV4 architecture which unified the ResNet and Inception based Spatio-temporal feature representation concept and achieving the highest top-1 accuracy till date. To extract and learn salient features for action recognition, attention driven residues are used which enhance the performance of residual components for effective 3D skeleton-based Spatio-temporal action representation. The robustness of the proposed framework is evaluated by performing extensive experiments on three challenging datasets such as UT Kinect Action 3D, Florence 3D action Dataset, and MSR Daily Action3D datasets, which consistently demonstrate the superiority of our method

📄 PDF Abstract BibTeX arXiv:1912.00576

Code (0)

등록된 구현이 없습니다.

Tasks

3D Action RecognitionAction RecognitionAction Recognition In VideosActivity RecognitionTemporal Action Localization

Methods 이 논문이 사용한 방법론

Average Pooling 설명 없음
ReLU How Do I Communicate to Expedia? How Do I Communicate to Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Live Support & Special Travel…
1x1 Convolution A 1 x 1 Convolution is a convolution with some special properties in that it can be used for dimensionality reduction,…
Batch Normalization 설명 없음
Bottleneck Residual Block A Bottleneck Residual Block is a variant of the residual block that utilises 1x1 convolutions to create a bottleneck. The…
Global Average Pooling Global Average Pooling is a pooling operation designed to replace fully connected layers in classical CNNs. The idea is to generate one feature map for each corresponding…
Residual Block Residual Blocks are skip-connection blocks that learn residual functions with reference to the layer inputs, instead of learning unreferenced functions. They were introduced…
Kaiming Initialization 설명 없음

Similar Papers 제목 키워드 기반

Expansion-Squeeze-Excitation Fusion Network for Elderly Activity Recognition

2021-12-21 · Xiangbo Shu, Jiawen Yang, Rui Yan, Yan Song

This work focuses on the task of elderly activity recognition, which is a challenging task due to the existence of individual actions and human-object interactions in elderly activities. Thus, we attempt to effectively a…

Action RecognitionActivity RecognitionHuman-Object Interaction Detection

GaitSTR: Gait Recognition with Sequential Two-stream Refinement

2024-04-02 · Wanrong Zheng, Haidong Zhu, Zhaoheng Zheng, Ram Nevatia

Gait recognition aims to identify a person based on their walking sequences, serving as a useful biometric modality as it can be observed from long distances without requiring cooperation from the subject. In representin…

Gait RecognitionMultiview Gait Recognition

SAFER-Activities: A Dataset for Smart Assessment of Fall Events and Routine Activities

2026-09-07 · Diwas Lamsal, Pramod Wickramatilake, Jednipat Moonrinta, Mongkol Ekpanyapong 외 arxiv

Smart healthcare monitoring systems require precise action recognition to ensure well-being and timely intervention in critical situations such as falls, particularly for mobility-challenged individuals. Existing dataset…

Action Recognition

Fusing Higher-order Features in Graph Neural Networks for Skeleton-based Action Recognition

2021-05-04 · Zhenyue Qin, Yang Liu, Pan Ji, Dongwoo Kim 외

Skeleton sequences are lightweight and compact, and thus are ideal candidates for action recognition on edge devices. Recent skeleton-based action recognition methods extract features from 3D joint coordinates as spatial…

Action RecognitionGraph Neural NetworkSkeleton Based Action Recognition

Three-Stream Convolutional Neural Network With Multi-Task and Ensemble Learning for 3D Action Recognition

2019-06-16 · The IEEE Conference on Computer Vision and Pattern Recognition (CVPR) Workshops, 2019 2019 6 · Duohan Liang, Guoliang Fan, Guangfeng Lin, Wanjun Chen 외

In this paper, we propose a three-stream convolutional neural network (3SCNN) for action recognition from skeleton sequences, which aims to thoroughly and fully exploit the skeleton data by extracting, learning, fusing a…

3D Action RecognitionAction RecognitionEnsemble LearningSkeleton Based Action Recognition