paper-with-me

Papers

FineParser: A Fine-grained Spatio-temporal Action Parser for Human-centric Action Quality Assessment

2024-05-11 · CVPR 2024 1 · Jinglin Xu, Sibo Yin, Guohao Zhao, Zishuo Wang, Yuxin Peng

Existing action quality assessment (AQA) methods mainly learn deep representations at the video level for scoring diverse actions. Due to the lack of a fine-grained understanding of actions in videos, they harshly suffer from low credibility and interpretability, thus insufficient for stringent applications, such as Olympic diving events. We argue that a fine-grained understanding of actions requires the model to perceive and parse actions in both time and space, which is also the key to the credibility and interpretability of the AQA technique. Based on this insight, we propose a new fine-grained spatial-temporal action parser named \textbf{FineParser}. It learns human-centric foreground action representations by focusing on target action regions within each frame and exploiting their fine-grained alignments in time and space to minimize the impact of invalid backgrounds during the assessment. In addition, we construct fine-grained annotations of human-centric foreground action masks for the FineDiving dataset, called \textbf{FineDiving-HM}. With refined annotations on diverse target action procedures, FineDiving-HM can promote the development of real-world AQA systems. Through extensive experiments, we demonstrate the effectiveness of FineParser, which outperforms state-of-the-art methods while supporting more tasks of fine-grained action understanding. Data and code are available at \url{https://github.com/PKU-ICST-MIPL/FineParser_CVPR2024}.

📄 PDF Abstract BibTeX arXiv:2405.06887

Code (1)

pku-icst-mipl/fineparser_cvpr2024 공식 구현 pytorch

Tasks

Action Quality AssessmentAction Understanding

Similar Papers 제목 키워드 기반

Decoupling Spatio-Temporal Adapter for Fine-Grained Badminton Action Localization

2026-05-22 · Tianyu Wang, Junjie Wu, Jingquan Gao, Shishuo Li arxiv

Temporal Action Localization (TAL) has been extensively studied in generic video understanding, while fine-grained sports scenarios, such as professional badminton, remain underexplored due to their complex and subtle sp…

Temporal Action Localization

Dynamic Spatio-Temporal Specialization Learning for Fine-Grained Action Recognition

2022-09-03 · Tianjiao Li, Lin Geng Foo, Qiuhong Ke, Hossein Rahmani 외

The goal of fine-grained action recognition is to successfully discriminate between action categories with subtle differences. To tackle this, we derive inspiration from the human visual system which contains specialized…

Action RecognitionFine-grained Action Recognition

Segmental Spatiotemporal CNNs for Fine-grained Action Segmentation

2016-02-09 · Colin Lea, Austin Reiter, Rene Vidal, Gregory D. Hager

Joint segmentation and classification of fine-grained actions is important for applications of human-robot interaction, video surveillance, and human skill evaluation. However, despite substantial recent progress in larg…

Action ClassificationAction RecognitionAction SegmentationFine-grained Action Recognition+3

FineMoGen: Fine-Grained Spatio-Temporal Motion Generation and Editing

2023-12-22 · NeurIPS 2023 11 · Mingyuan Zhang, Huirong Li, Zhongang Cai, Jiawei Ren 외

Text-driven motion generation has achieved substantial progress with the emergence of diffusion models. However, existing methods still struggle to generate complex motion sequences that correspond to fine-grained descri…

Mixture-of-ExpertsMotion GenerationMotion Synthesis

Spatio-temporal Gait Feature with Global Distance Alignment

2022-03-07 · Yifan Chen, Yang Zhao, Xuelong Li

Gait recognition is an important recognition technology, because gait is not easy to camouflage and does not need cooperation to recognize subjects. However, many existing methods are inadequate in preserving both tempor…

Gait Recognition