paper-with-me

홈 › Papers

Temporal Feature Alignment and Mutual Information Maximization for Video-Based Human Pose Estimation

2022-03-29 · CVPR 2022 1 · Zhenguang Liu, Runyang Feng, Haoming Chen, Shuang Wu, Yixing Gao, Yunjun Gao, Xiang Wang

Multi-frame human pose estimation has long been a compelling and fundamental problem in computer vision. This task is challenging due to fast motion and pose occlusion that frequently occur in videos. State-of-the-art methods strive to incorporate additional visual evidences from neighboring frames (supporting frames) to facilitate the pose estimation of the current frame (key frame). One aspect that has been obviated so far, is the fact that current methods directly aggregate unaligned contexts across frames. The spatial-misalignment between pose features of the current frame and neighboring frames might lead to unsatisfactory results. More importantly, existing approaches build upon the straightforward pose estimation loss, which unfortunately cannot constrain the network to fully leverage useful information from neighboring frames. To tackle these problems, we present a novel hierarchical alignment framework, which leverages coarse-to-fine deformations to progressively update a neighboring frame to align with the current frame at the feature level. We further propose to explicitly supervise the knowledge extraction from neighboring frames, guaranteeing that useful complementary cues are extracted. To achieve this goal, we theoretically analyzed the mutual information between the frames and arrived at a loss that maximizes the task-relevant mutual information. These allow us to rank No.1 in the Multi-frame Person Pose Estimation Challenge on benchmark dataset PoseTrack2017, and obtain state-of-the-art performance on benchmarks Sub-JHMDB and Pose-Track2018. Our code is released at https://github. com/Pose-Group/FAMI-Pose, hoping that it will be useful to the community.

📄 PDF Abstract BibTeX arXiv:2203.15227

Code (1)

pose-group/fami-pose 공식 구현 pytorch

Tasks

Pose Estimation

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

Zero-shot Skeleton-based Action Recognition via Mutual Information Estimation and Maximization

2023-08-07 · Yujie Zhou, Wenwen Qiang, Anyi Rao, Ning Lin 외

Zero-shot skeleton-based action recognition aims to recognize actions of unseen categories after training on data of seen categories. The key is to build the connection between visual and semantic space from seen to unse…

Action RecognitionMutual Information EstimationSkeleton Based Action RecognitionZero Shot Skeletal Action Recognition+1

Few-shot Action Recognition via Intra- and Inter-Video Information Maximization

2023-05-10 · Huabin Liu, Weiyao Lin, Tieyuan Chen, Yuxi Li 외

Current few-shot action recognition involves two primary sources of information for classification:(1) intra-video information, determined by frame content within a single video clip, and (2) inter-video information, mea…

Action RecognitionFew-Shot action recognitionFew Shot Action RecognitionTemporal Action Localization+1

Diverse Melody Generation from Chinese Lyrics via Mutual Information Maximization

2020-12-07 · Ruibin Yuan, Ge Zhang, Anqiao Yang, Xinyue Zhang

In this paper, we propose to adapt the method of mutual information maximization into the task of Chinese lyrics conditioned melody generation to improve the generation quality and diversity. We employ scheduled sampling…

Diversity

Spatio-Temporal Deep Graph Infomax

2019-04-12 · Felix L. Opolka, Aaron Solomon, Cătălina Cangea, Petar Veličković 외

Spatio-temporal graphs such as traffic networks or gene regulatory systems present challenges for the existing deep learning methods due to the complexity of structural changes over time. To address these issues, we intr…

Representation LearningTraffic Prediction

High-Order Conditional Mutual Information Maximization for dealing with High-Order Dependencies in Feature Selection

2022-07-18 · Francisco Souza, Cristiano Premebida, Rui Araújo

This paper presents a novel feature selection method based on the conditional mutual information (CMI). The proposed High Order Conditional Mutual Information Maximization (HOCMIM) incorporates high order dependencies in…

feature selectionVocal Bursts Intensity Prediction