paper-with-me

Papers

Body Joint guided 3D Deep Convolutional Descriptors for Action Recognition

2017-04-24 · Congqi Cao, Yifan Zhang, Chunjie Zhang, Hanqing Lu

Three dimensional convolutional neural networks (3D CNNs) have been established as a powerful tool to simultaneously learn features from both spatial and temporal dimensions, which is suitable to be applied to video-based action recognition. In this work, we propose not to directly use the activations of fully-connected layers of a 3D CNN as the video feature, but to use selective convolutional layer activations to form a discriminative descriptor for video. It pools the feature on the convolutional layers under the guidance of body joint positions. Two schemes of mapping body joints into convolutional feature maps for pooling are discussed. The body joint positions can be obtained from any off-the-shelf skeleton estimation algorithm. The helpfulness of the body joint guided feature pooling with inaccurate skeleton estimation is systematically evaluated. To make it end-to-end and do not rely on any sophisticated body joint detection algorithm, we further propose a two-stream bilinear model which can learn the guidance from the body joints and capture the spatio-temporal features simultaneously. In this model, the body joint guided feature pooling is conveniently formulated as a bilinear product operation. Experimental results on three real-world datasets demonstrate the effectiveness of body joint guided pooling which achieves promising performance.

📄 PDF Abstract BibTeX arXiv:1704.07160

Code (0)

등록된 구현이 없습니다.

Tasks

Action RecognitionTemporal Action Localization

Similar Papers 제목 키워드 기반

A Semantics-Guided Graph Convolutional Network for Skeleton-Based Action Recognition

2020-05-01 · ICIAI 2020: Proceedings of the 2020 the 4th International Conference on Innovation in Artificial Intelligence 2020 5 · Xiaolu Ding, Kai Yang, Wai Chen

Action recognition with skeleton data is a challenging task in computer vision. Graph convolutional networks (GCNs), which directly model the human body skeletons as the graph structure, have achieved remarkable performa…

Action RecognitionSkeleton Based Action Recognition

Part-Aligned Bilinear Representations for Person Re-identification

2018-04-19 · ECCV 2018 9 · Yumin Suh, Jingdong Wang, Siyu Tang, Tao Mei 외

We propose a novel network that learns a part-aligned representation for person re-identification. It handles the body part misalignment problem, that is, body parts are misaligned across human detections due to pose/vie…

2D Human Pose EstimationPerson Re-IdentificationPose Estimation

A Human Action Descriptor Based on Motion Coordination

2019-11-20 · Pietro Falco, Matteo Saveriano, Eka Gibran Hasany, Nicholas H. Kirk 외

In this paper, we present a descriptor for human whole-body actions based on motion coordination. We exploit the principle, well known in neuromechanics, that humans move their joints in a coordinated fashion. Our coordi…

Probabilistic Human Mesh Recovery in 3D Scenes from Egocentric Views

2023-04-12 · ICCV 2023 1 · Siwei Zhang, Qianli Ma, Yan Zhang, Sadegh Aliakbarian 외

Automatic perception of human behaviors during social interactions is crucial for AR/VR applications, and an essential component is estimation of plausible 3D human pose and shape of our social partners from the egocentr…

DiversityHuman Mesh Recovery

mmJoints: Expanding Joint Representations Beyond (x,y,z) in mmWave-Based 3D Pose Estimation

2025-10-10 · Zhenyu Wang, Mahathir Monjur, Shahriar Nirjon arxiv

In mmWave-based pose estimation, sparse signals and weak reflections often cause models to infer body joints from statistical priors rather than sensor data. While prior knowledge helps in learning meaningful representat…

Activity Recognition3D Pose Estimation