paper-with-me

Papers

Action Segmentation Using 2D Skeleton Heatmaps and Multi-Modality Fusion

2023-09-12 · Syed Waleed Hyder, Muhammad Usama, Anas Zafar, Muhammad Naufil, Fawad Javed Fateh, Andrey Konin, M. Zeeshan Zia, Quoc-Huy Tran

This paper presents a 2D skeleton-based action segmentation method with applications in fine-grained human activity recognition. In contrast with state-of-the-art methods which directly take sequences of 3D skeleton coordinates as inputs and apply Graph Convolutional Networks (GCNs) for spatiotemporal feature learning, our main idea is to use sequences of 2D skeleton heatmaps as inputs and employ Temporal Convolutional Networks (TCNs) to extract spatiotemporal features. Despite lacking 3D information, our approach yields comparable/superior performances and better robustness against missing keypoints than previous methods on action segmentation datasets. Moreover, we improve the performances further by using both 2D skeleton heatmaps and RGB videos as inputs. To our best knowledge, this is the first work to utilize 2D skeleton heatmap inputs and the first work to explore 2D skeleton+RGB fusion for action segmentation.

📄 PDF Abstract BibTeX arXiv:2309.06462

Code (0)

등록된 구현이 없습니다.

Tasks

Action SegmentationActivity RecognitionHuman Activity RecognitionSegmentationSkeleton Based Action Segmentation

Methods 이 논문이 사용한 방법론

Heatmap 설명 없음

Similar Papers 제목 키워드 기반

Learning by Aligning 2D Skeleton Sequences and Multi-Modality Fusion

2023-05-31 · Quoc-Huy Tran, Muhammad Ahmed, Murad Popattia, M. Hassan Ahmed 외

This paper presents a self-supervised temporal video alignment framework which is useful for several fine-grained human activity understanding applications. In contrast with the state-of-the-art method of CASA, where seq…

RetrievalSelf-Supervised LearningVideo Alignment

Fine-grained Action Analysis: A Multi-modality and Multi-task Dataset of Figure Skating

2023-07-06 · Sheng-Lan Liu, Yu-Ning Ding, Gang Yan, Si-Fan Zhang 외

The fine-grained action analysis of the existing action datasets is challenged by insufficient action categories, low fine granularities, limited modalities, and tasks. In this paper, we propose a Multi-modality and Mult…

Action AnalysisAction Quality AssessmentAction RecognitionFine-grained Action Recognition

Language-Assisted Skeleton Action Understanding for Skeleton-Based Temporal Action Segmentation

2024-10-31 · European Conference on Computer Vision (ECCV2024) 2024 10 · Haoyu Ji, Bowen Chen, Xinglong Xu, Weihong Ren 외

Skeleton-based Temporal Action Segmentation (STAS) aims to densely segment and classify human actions in long, untrimmed skeletal motion sequences. Existing STAS methods primarily model spatial dependencies among joints …

Action SegmentationAction UnderstandingContrastive LearningRepresentation Learning+4

EV-Action: Electromyography-Vision Multi-Modal Action Dataset

2019-04-20 · Lichen Wang, Bin Sun, Joseph Robinson, Taotao Jing 외

Multi-modal human action analysis is a critical and attractive research topic. However, the majority of the existing datasets only provide visual modalities (i.e., RGB, depth and skeleton). To make up this, we introduce …

Action AnalysisAction RecognitionElectromyography (EMG)Multimodal Activity Recognition+1

Unified Multi-modal Unsupervised Representation Learning for Skeleton-based Action Understanding

2023-11-06 · Shengkai Sun, Daizong Liu, Jianfeng Dong, Xiaoye Qu 외

Unsupervised pre-training has shown great success in skeleton-based action understanding recently. Existing works typically train separate modality-specific models, then integrate the multi-modal information for action u…

Action UnderstandingRepresentation LearningUnsupervised Pre-training