paper-with-me

홈 › Papers

Towards Efficient General Feature Prediction in Masked Skeleton Modeling

2025-09-03 · Shengkai Sun, Zefan Zhang, Jianfeng Dong, Zhiyong Cheng, Xiaojun Chang, Meng Wang arxiv

Recent advances in the masked autoencoder (MAE) paradigm have significantly propelled self-supervised skeleton-based action recognition. However, most existing approaches limit reconstruction targets to raw joint coordinates or their simple variants, resulting in computational redundancy and limited semantic representation. To address this, we propose a novel General Feature Prediction framework (GFP) for efficient mask skeleton modeling. Our key innovation is replacing conventional low-level reconstruction with high-level feature prediction that spans from local motion patterns to global semantic representations. Specifically, we introduce a collaborative learning framework where a lightweight target generation network dynamically produces diversified supervision signals across spatial-temporal hierarchies, avoiding reliance on pre-computed offline features. The framework incorporates constrained optimization to ensure feature diversity while preventing model collapse. Experiments on NTU RGB+D 60, NTU RGB+D 120 and PKU-MMD demonstrate the benefits of our approach: Computational efficiency (with 6.2$\times$ faster training than standard masked skeleton modeling methods) and superior representation quality, achieving state-of-the-art performance in various downstream tasks.

📄 PDF Abstract BibTeX arXiv:2509.03609

Code (0)

등록된 구현이 없습니다.

Tasks

Computational EfficiencyAction Recognition

Similar Papers 제목 키워드 기반

Prompted Contrast with Masked Motion Modeling: Towards Versatile 3D Action Representation Learning

2023-08-08 · Jiahang Zhang, Lilang Lin, Jiaying Liu

Self-supervised learning has proved effective for skeleton-based human action understanding, which is an important yet challenging topic. Previous works mainly rely on contrastive learning or masked motion modeling parad…

Action UnderstandingContrastive LearningRepresentation LearningSelf-Supervised Learning

SimMC: Simple Masked Contrastive Learning of Skeleton Representations for Unsupervised Person Re-Identification

2022-04-21 · Haocong Rao, Chunyan Miao

Recent advances in skeleton-based person re-identification (re-ID) obtain impressive performance via either hand-crafted skeleton descriptors or skeleton representation learning with deep learning paradigms. However, the…

Contrastive LearningPerson Re-IdentificationRepresentation LearningUnsupervised Person Re-Identification

SkeletonMAE: Spatial-Temporal Masked Autoencoders for Self-supervised Skeleton Action Recognition

2022-09-01 · Wenhan Wu, Yilei Hua, Ce Zheng, Shiqian Wu 외

Fully supervised skeleton-based action recognition has achieved great progress with the blooming of deep learning techniques. However, these methods require sufficient labeled data which is not easy to obtain. In contras…

Action RecognitionDecoderSelf-supervised Skeleton-based Action RecognitionSkeleton Based Action Recognition

Masked Motion Predictors are Strong 3D Action Representation Learners

2023-08-14 · ICCV 2023 1 · Yunyao Mao, Jiajun Deng, Wengang Zhou, Yao Fang 외

In 3D human action recognition, limited supervised data makes it challenging to fully tap into the modeling potential of powerful networks such as transformers. As a result, researchers have been actively investigating e…

3D Action RecognitionAction RecognitionFew-Shot Skeleton-Based Action Recognitionmotion prediction+3

Skeleton2vec: A Self-supervised Learning Framework with Contextualized Target Representations for Skeleton Sequence

2024-01-01 · Ruizhuo Xu, Linzhi Huang, Mei Wang, Jiani Hu 외

Self-supervised pre-training paradigms have been extensively explored in the field of skeleton-based action recognition. In particular, methods based on masked prediction have pushed the performance of pre-training to a …

Action RecognitionPredictionRepresentation LearningSelf-Supervised Learning+1