paper-with-me

홈 › Papers

Pyramid Self-attention Polymerization Learning for Semi-supervised Skeleton-based Action Recognition

2023-02-05 · Binqian Xu, Xiangbo Shu

Most semi-supervised skeleton-based action recognition approaches aim to learn the skeleton action representations only at the joint level, but neglect the crucial motion characteristics at the coarser-grained body (e.g., limb, trunk) level that provide rich additional semantic information, though the number of labeled data is limited. In this work, we propose a novel Pyramid Self-attention Polymerization Learning (dubbed as PSP Learning) framework to jointly learn body-level, part-level, and joint-level action representations of joint and motion data containing abundant and complementary semantic information via contrastive learning covering coarse-to-fine granularity. Specifically, to complement semantic information from coarse to fine granularity in skeleton actions, we design a new Pyramid Polymerizing Attention (PPA) mechanism that firstly calculates the body-level attention map, part-level attention map, and joint-level attention map, as well as polymerizes these attention maps in a level-by-level way (i.e., from body level to part level, and further to joint level). Moreover, we present a new Coarse-to-fine Contrastive Loss (CCL) including body-level contrast loss, part-level contrast loss, and joint-level contrast loss to jointly measure the similarity between the body/part/joint-level contrasting features of joint and motion data. Finally, extensive experiments are conducted on the NTU RGB+D and North-Western UCLA datasets to demonstrate the competitive performance of the proposed PSP Learning in the semi-supervised skeleton-based action recognition task. The source codes of PSP Learning are publicly available at https://github.com/1xbq1/PSP-Learning.

📄 PDF Abstract BibTeX arXiv:2302.02327

Code (1)

1xbq1/psp-learning 공식 구현 pytorch

Tasks

Action RecognitionContrastive LearningSkeleton Based Action Recognition

Methods 이 논문이 사용한 방법론

Contrastive Learning 설명 없음

Similar Papers 제목 키워드 기반

Cross-pyramid consistency regularization for semi-supervised medical image segmentation

2025-11-11 · Matus Bojko, Maros Kollar, Marek Jakab, Wanda Benesova arxiv

Semi-supervised learning (SSL) enables training of powerful models with the assumption of limited, carefully labelled data and a large amount of unlabeled data to support the learning. In this paper, we propose a hybrid …

Semi-supervised Medical Image SegmentationSelf-Supervised LearningKnowledge Distillation

AstMatch: Adversarial Self-training Consistency Framework for Semi-Supervised Medical Image Segmentation

2024-06-28 · Guanghao Zhu, Jing Zhang, Juanxiu Liu, Xiaohui Du 외

Semi-supervised learning (SSL) has shown considerable potential in medical image segmentation, primarily leveraging consistency regularization and pseudo-labeling. However, many SSL approaches only pay attention to low-l…

Image SegmentationMedical Image SegmentationPseudo LabelSemantic Segmentation+2

PVStereo: Pyramid Voting Module for End-to-End Self-Supervised Stereo Matching

2021-03-12 · Hengli Wang, Rui Fan, Peide Cai, Ming Liu

Supervised learning with deep convolutional neural networks (DCNNs) has seen huge adoption in stereo matching. However, the acquisition of large-scale datasets with well-labeled ground truth is cumbersome and labor-inten…

Stereo Matching

TreeFormer: a Semi-Supervised Transformer-based Framework for Tree Counting from a Single High Resolution Image

2023-07-12 · Hamed Amini Amirkolaee, Miaojing Shi, Mark Mulligan

Automatic tree density estimation and counting using single aerial and satellite images is a challenging task in photogrammetry and remote sensing, yet has an important role in forest management. In this paper, we propos…

DecoderDensity Estimation

Incorporating Temporal Prior from Motion Flow for Instrument Segmentation in Minimally Invasive Surgery Video

2019-07-18 · Yueming Jin, Keyun Cheng, Qi Dou, Pheng-Ann Heng

Automatic instrument segmentation in video is an essentially fundamental yet challenging problem for robot-assisted minimally invasive surgery. In this paper, we propose a novel framework to leverage instrument motion in…

DecoderSegmentation