paper-with-me

홈 › Papers

Infrared and 3D skeleton feature fusion for RGB-D action recognition

2020-02-28 · submitted to IEEE Access 2020 2 · Alban Main de Boissiere, Rita Noumeir

A challenge of skeleton-based action recognition is the difficulty to classify actions with similar motions and object-related actions. Visual clues from other streams help in that regard. RGB data are sensible to illumination conditions, thus unusable in the dark. To alleviate this issue and still benefit from a visual stream, we propose a modular network (FUSION) combining skeleton and infrared data. A 2D convolutional neural network (CNN) is used as a pose module to extract features from skeleton data. A 3D CNN is used as an infrared module to extract visual cues from videos. Both feature vectors are then concatenated and exploited conjointly using a multilayer perceptron (MLP). Skeleton data also condition the infrared videos, providing a crop around the performing subjects and thus virtually focusing the attention of the infrared module. Ablation studies show that using pre-trained networks on other large scale datasets as our modules and data augmentation yield considerable improvements on the action classification accuracy. The strong contribution of our cropping strategy is also demonstrated. We evaluate our method on the NTU RGB+D dataset, the largest dataset for human action recognition from depth cameras, and report state-of-the-art performances.

📄 PDF Abstract BibTeX arXiv:2002.12886

Code (1)

adeboissiere/FUSION-human-action-recognition 공식 구현 pytorch

Tasks

Action ClassificationAction RecognitionData AugmentationSkeleton Based Action RecognitionTemporal Action Localization

Similar Papers 제목 키워드 기반

TDSM: Triplet Diffusion for Skeleton-Text Matching in Zero-Shot Action Recognition

2024-11-16 · Jeonghyeok Do, Munchurl Kim

We firstly present a diffusion-based action recognition with zero-shot learning for skeleton inputs. In zero-shot skeleton-based action recognition, aligning skeleton features with the text features of action labels is e…

Action RecognitionSkeleton Based Action RecognitionText MatchingTriplet+3

Learning Rich Features for Gait Recognition by Integrating Skeletons and Silhouettes

2021-10-26 · Yunjie Peng, Kang Ma, Yang Zhang, Zhiqiang He

Gait recognition captures gait patterns from the walking sequence of an individual for identification. Most existing gait recognition methods learn features from silhouettes or skeletons for the robustness to clothing, c…

Gait IdentificationGait Recognition

Fusing Higher-order Features in Graph Neural Networks for Skeleton-based Action Recognition

2021-05-04 · Zhenyue Qin, Yang Liu, Pan Ji, Dongwoo Kim 외

Skeleton sequences are lightweight and compact, and thus are ideal candidates for action recognition on edge devices. Recent skeleton-based action recognition methods extract features from 3D joint coordinates as spatial…

Action RecognitionGraph Neural NetworkSkeleton Based Action Recognition

Skeleton Sequence and RGB Frame Based Multi-Modality Feature Fusion Network for Action Recognition

2022-02-23 · Xiaoguang Zhu, Ye Zhu, Haoyu Wang, Honglin Wen 외

Action recognition has been a heated topic in computer vision for its wide application in vision systems. Previous approaches achieve improvement by fusing the modalities of the skeleton sequence and RGB video. However, …

Action Recognition

Pose-Guided Graph Convolutional Networks for Skeleton-Based Action Recognition

2022-10-10 · Han Chen, Yifan Jiang, Hanseok Ko

Graph convolutional networks (GCNs), which can model the human body skeletons as spatial and temporal graphs, have shown remarkable potential in skeleton-based action recognition. However, in the existing GCN-based metho…

Action RecognitionSkeleton Based Action RecognitionTemporal Action Localization