paper-with-me

홈 › Papers

Temporal Keypoint Matching and Refinement Network for Pose Estimation and Tracking

2020-08-01 · ECCV 2020 8 · Chunluan Zhou Zhou Ren Gang Hua

Multi-person pose estimation and tracking in realistic videos is very challenging due to factors such as occlusions, fast motion and pose variations. Top-down approaches are commonly used for this task, which involves three stages: person detection, single-person pose estimation, and pose association across time. Recently, significant progress has been made in person detection and single-person pose estimation. In this paper, we mainly focus on improving pose association and estimation in a video to build a strong pose estimator and tracker. To this end, we propose a novel temporal keypoint matching and refinement network. Specifically, we propose two network modules, temporal keypoint matching and temporal keypoint refinement, which are incorporated into a single-person pose estimatin network. The temporal keypoint matching module learns a simialrity metric for matching keypoints across frames. Pose matching is performed by aggregating keypoint similarities between poses in adjacent frames. The temporal keypoint refinement module serves to correct individual poses by utilizing their associated poses in neighboring frames as temporal context. We validate the effectiveness of our proposed network on two benchmark datasets: PoseTrack 2017 and PoseTrack 2018. Exprimental results show that our approach achieves state-of-the-art performance on both datasets.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Human DetectionMulti-Person Pose EstimationMulti-Person Pose Estimation and TrackingPose Estimation

Similar Papers 제목 키워드 기반

15 Keypoints Is All You Need

2019-12-05 · CVPR 2020 6 · Michael Snower, Asim Kadav, Farley Lai, Hans Peter Graf

Pose tracking is an important problem that requires identifying unique human pose-instances and matching them temporally across different frames of a video. However, existing pose tracking methods are unable to accuratel…

AllBinary ClassificationOptical Flow EstimationPose Estimation+1

XRefine: Attention-Guided Keypoint Match Refinement

2026-01-18 · Jan Fabian Schmid, Annika Hagemann arxiv

Sparse keypoint matching is crucial for 3D vision tasks, yet current keypoint detectors often produce spatially inaccurate matches. Existing refinement methods mitigate this issue through alignment of matched keypoint lo…

Matching Is Not Enough: A Two-Stage Framework for Category-Agnostic Pose Estimation

2023-01-01 · CVPR 2023 1 · Min Shi, Zihao Huang, Xianzheng Ma, Xiaowei Hu 외

Category-agnostic pose estimation (CAPE) aims to predict keypoints for arbitrary categories given support images with keypoint annotations. Existing approaches match the keypoints across the image for localization. H…

2D Pose EstimationCategory-Agnostic Pose EstimationDecoderPose Estimation

Joint angle model based learning to refine kinematic human pose estimation

2025-07-15 · Chang Peng, Yifei Zhou, Huifeng Xi, Shiqing Huang 외

Marker-free human pose estimation (HPE) has found increasing applications in various fields. Current HPE suffers from occasional errors in keypoint recognition and random fluctuation in keypoint trajectories when analyzi…

Pose Estimation

Video Action Segmentation via Contextually Refined Temporal Keypoints

2023-01-01 · ICCV 2023 1 · Borui Jiang, Yang Jin, Zhentao Tan, Yadong Mu

Video action segmentation refers to the task of densely casting each video frame or short segment in an untrimmed video into some pre-specified action categories. Although recent years have witnessed a great promise …

Action SegmentationGraph MatchingSegmentation