Unsupervised Human Pose Estimation through Transforming Shape Templates
Human pose estimation is a major computer vision problem with applications ranging from augmented reality and video capture to surveillance and movement tracking. In the medical context, the latter may be an important biomarker for neurological impairments in infants. Whilst many methods exist, their application has been limited by the need for well annotated large datasets and the inability to generalize to humans of different shapes and body compositions, e.g. children and infants. In this paper we present a novel method for learning pose estimators for human adults and infants in an unsupervised fashion. We approach this as a learnable template matching problem facilitated by deep feature extractors. Human-interpretable landmarks are estimated by transforming a template consisting of predefined body parts that are characterized by 2D Gaussian distributions. Enforcing a connectivity prior guides our model to meaningful human shape representations. We demonstrate the effectiveness of our approach on two different datasets including adults and infants.
Code (2)
Tasks
Pose EstimationTemplate MatchingUnsupervised Human Pose EstimationSimilar Papers 제목 키워드 기반
The Edge of Depth: Explicit Constraints between Segmentation and Depth
In this work we study the mutual benefits of two common computer vision tasks, self-supervised depth estimation and semantic segmentation from images. For example, to help unsupervised monocular depth estimation, constra…
Depth EstimationMonocular Depth EstimationSegmentationSemantic Segmentation+1Unsupervised Domain Adaptation with Copula Models
We study the task of unsupervised domain adaptation, where no labeled data from the target domain is provided during training time. To deal with the potential discrepancy between the source and target distributions, both…
Domain AdaptationregressionUnsupervised Domain AdaptationBack to the Color: Learning Depth to Specific Color Transformation for Unsupervised Depth Estimation
Virtual engines can generate dense depth maps for various synthetic scenes, making them invaluable for training depth estimation models. However, discrepancies between synthetic and real-world colors pose significant cha…
Depth EstimationMonocular Depth EstimationUnsupervised Monocular Depth EstimationX as Supervision: Contending with Depth Ambiguity in Unsupervised Monocular 3D Pose Estimation
Recent unsupervised methods for monocular 3D pose estimation have endeavored to reduce dependence on limited annotated 3D data, but most are solely formulated in 2D space, overlooking the inherent depth ambiguity issue. …
3D Pose EstimationPose EstimationJoint COCO and Mapillary Workshop at ICCV 2019 Keypoint Detection Challenge Track Technical Report: Distribution-Aware Coordinate Representation for Human Pose Estimation
In this paper, we focus on the coordinate representation in human pose estimation. While being the standard choice, heatmap based representation has not been systematically investigated. We found that the process of coor…
Keypoint DetectionPose Estimation