paper-with-me

Papers

Self-Supervised Skeleton-Based Action Representation Learning: A Benchmark and Beyond

2024-06-05 · Jiahang Zhang, Lilang Lin, Shuai Yang, Jiaying Liu

Self-supervised learning (SSL), which aims to learn meaningful prior representations from unlabeled data, has been proven effective for skeleton-based action understanding. Different from the image domain, skeleton data possesses sparser spatial structures and diverse representation forms, with the absence of background clues and the additional temporal dimension, presenting new challenges for spatial-temporal motion pretext task design. Recently, many endeavors have been made for skeleton-based SSL, achieving remarkable progress. However, a systematic and thorough review is still lacking. In this paper, we conduct, for the first time, a comprehensive survey on self-supervised skeleton-based action representation learning. Following the taxonomy of context-based, generative learning, and contrastive learning approaches, we make a thorough review and benchmark of existing works and shed light on the future possible directions. Remarkably, our investigation demonstrates that most SSL works rely on the single paradigm, learning representations of a single level, and are evaluated on the action recognition task solely, which leaves the generalization power of skeleton SSL models under-explored. To this end, a novel and effective SSL method for skeleton is further proposed, which integrates versatile representation learning objectives of different granularity, substantially boosting the generalization capacity for multiple skeleton downstream tasks. Extensive experiments under three large-scale datasets demonstrate our method achieves superior generalization performance on various downstream tasks, including recognition, retrieval, detection, and few-shot learning.

📄 PDF Abstract BibTeX arXiv:2406.02978

Code (1)

jhang2020/pcm3 공식 구현 pytorch

Tasks

Action RecognitionAction UnderstandingContrastive LearningFew-Shot LearningRepresentation LearningSelf-Supervised Learning

Methods 이 논문이 사용한 방법론

Contrastive Learning 설명 없음

Similar Papers 제목 키워드 기반

Self-Supervised 3D Action Representation Learning with Skeleton Cloud Colorization

2023-04-18 · Siyuan Yang, Jun Liu, Shijian Lu, Er Meng Hwa 외

3D Skeleton-based human action recognition has attracted increasing attention in recent years. Most of the existing work focuses on supervised learning which requires a large number of labeled action sequences that are o…

3D Action RecognitionAction RecognitionColorizationRepresentation Learning+3

Skeleton-Contrastive 3D Action Representation Learning

2021-08-08 · Fida Mohammad Thoker, Hazel Doughty, Cees G. M. Snoek

This paper strives for self-supervised learning of a feature space suitable for skeleton-based action recognition. Our proposal is built upon learning invariances to input skeleton representations and various skeleton au…

Action RecognitionContrastive LearningFew-Shot Skeleton-Based Action RecognitionRepresentation Learning+4

ViA: View-invariant Skeleton Action Representation Learning via Motion Retargeting

2022-08-31 · Di Yang, Yaohui Wang, Antitza Dantcheva, Lorenzo Garattoni 외

Current self-supervised approaches for skeleton action representation learning often focus on constrained scenarios, where videos and skeleton data are recorded in laboratory settings. When dealing with estimated skeleto…

Action ClassificationAction Recognitionmotion retargetingRepresentation Learning+2

Bootstrapped Representation Learning for Skeleton-Based Action Recognition

2022-02-04 · Olivier Moliner, Sangxia Huang, Kalle Åström

In this work, we study self-supervised representation learning for 3D skeleton-based action recognition. We extend Bootstrap Your Own Latent (BYOL) for representation learning on skeleton sequence data and propose a new …

Action RecognitionData AugmentationKnowledge DistillationLinear evaluation+3

Hierarchically Self-Supervised Transformer for Human Skeleton Representation Learning

2022-07-20 · Yuxiao Chen, Long Zhao, Jianbo Yuan, Yu Tian 외

Despite the success of fully-supervised human skeleton sequence modeling, utilizing self-supervised pre-training for skeleton sequence representation learning has been an active field because acquiring task-specific skel…

Action DetectionAction RecognitionContrastive Learningmotion prediction+1