paper-with-me

Papers

A Baseline Framework for Part-level Action Parsing and Action Recognition

2021-10-07 · Xiaodong Chen, Xinchen Liu, Kun Liu, Wu Liu, Tao Mei

This technical report introduces our 2nd place solution to Kinetics-TPS Track on Part-level Action Parsing in ICCV DeeperAction Workshop 2021. Our entry is mainly based on YOLOF for instance and part detection, HRNet for human pose estimation, and CSN for video-level action recognition and frame-level part state parsing. We describe technical details for the Kinetics-TPS dataset, together with some experimental results. In the competition, we achieved 61.37% mAP on the test set of Kinetics-TPS.

📄 PDF Abstract BibTeX arXiv:2110.03368

Code (0)

등록된 구현이 없습니다.

Tasks

Action ParsingAction RecognitionPose Estimation

Methods 이 논문이 사용한 방법론

Test 설명 없음
ReLU How Do I Communicate to Expedia? How Do I Communicate to Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Live Support & Special Travel…
Residual Connection 설명 없음
Batch Normalization 설명 없음
Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…
HRNet HRNet, or High-Resolution Net, is a general purpose convolutional neural network for tasks like semantic segmentation, object detection and image classification. It is…

Similar Papers 제목 키워드 기반

Technical Report: Disentangled Action Parsing Networks for Accurate Part-level Action Parsing

2021-11-05 · Xuanhan Wang, Xiaojia Chen, Lianli Gao, Lechao Chen 외

Part-level Action Parsing aims at part state parsing for boosting action recognition in videos. Despite of dramatic progresses in the area of video classification research, a severe problem faced by the community is that…

Action ParsingAction RecognitionAction Recognition In VideosHuman Detection+1

Part-level Action Parsing via a Pose-guided Coarse-to-Fine Framework

2022-03-09 · Xiaodong Chen, Xinchen Liu, Wu Liu, Kun Liu 외

Action recognition from videos, i.e., classifying a video into one of the pre-defined action types, has been a popular topic in the communities of artificial intelligence, multimedia, and signal processing. However, exis…

Action ParsingAction Recognition

Part-aware Panoptic Segmentation

2021-06-11 · CVPR 2021 1 · Daan de Geus, Panagiotis Meletis, Chenyang Lu, Xiaoxiao Wen 외

In this work, we introduce the new scene understanding task of Part-aware Panoptic Segmentation (PPS), which aims to understand a scene at multiple levels of abstraction, and unifies the tasks of scene parsing and part p…

Image SegmentationPanoptic SegmentationPart-aware Panoptic SegmentationScene Parsing+2

Logics-Parsing-Omni Technical Report

2026-03-10 · Xin An, Jingyi Cai, Xiangyang Chen, Huayao Liu 외 arxiv

Addressing the challenges of fragmented task definitions and the heterogeneity of unstructured data in multimodal parsing, this paper proposes the Omni Parsing framework. This framework establishes a Unified Taxonomy cov…

Attribute Extraction

Multi-Class Part Parsing With Joint Boundary-Semantic Awareness

2019-10-01 · ICCV 2019 10 · Yifan Zhao, Jia Li, Yu Zhang, Yonghong Tian

Object part parsing in the wild, which requires to simultaneously detect multiple object classes in the scene and accurately segments semantic parts within each class, is challenging for the joint presence of class-level…

2D Semantic Segmentation