SVMAC: Unsupervised 3D Human Pose Estimation from a Single Image with Single-view-multi-angle Consistency
Recovering 3D human pose from 2D joints is still a challenging problem, especially without any 3D annotation, video information, or multi-view information. In this paper, we present an unsupervised GAN-based model consisting of multiple weight-sharing generators to estimate a 3D human pose from a single image without 3D annotations. In our model, we introduce single-view-multi-angle consistency (SVMAC) to significantly improve the estimation performance. With 2D joint locations as input, our model estimates a 3D pose and a camera simultaneously. During training, the estimated 3D pose is rotated by random angles and the estimated camera projects the rotated 3D poses back to 2D. The 2D reprojections will be fed into weight-sharing generators to estimate the corresponding 3D poses and cameras, which are then mixed to impose SVMAC constraints to self-supervise the training process. The experimental results show that our method outperforms the state-of-the-art unsupervised methods on Human 3.6M and MPI-INF-3DHP. Moreover, qualitative results on MPII and LSP show that our method can generalize well to unknown data.
Code (0)
등록된 구현이 없습니다.
Tasks
3D Human Pose EstimationMonocular 3D Human Pose EstimationPose EstimationUnsupervised 3D Human Pose EstimationSimilar Papers 제목 키워드 기반
ElePose: Unsupervised 3D Human Pose Estimation by Predicting Camera Elevation and Learning Normalizing Flows on 2D Poses
Human pose estimation from single images is a challenging problem that is typically solved by supervised learning. Unfortunately, labeled training data does not yet exist for many human activities since 3D annotation req…
3D Human Pose EstimationPose EstimationUnsupervised 3D Human Pose EstimationUnsupervised Human Pose EstimationSharinGAN: Combining Synthetic and Real Data for Unsupervised Geometry Estimation
We propose a novel method for combining synthetic and real images when training networks to determine geometric information from a single image. We suggest a method for mapping both image types into a single, shared doma…
Depth EstimationMonocular Depth EstimationSurface Normal EstimationSurface Normals Estimation+1Unsupervised Adversarial Learning of 3D Human Pose from 2D Joint Locations
The task of three-dimensional (3D) human pose estimation from a single image can be divided into two parts: (1) Two-dimensional (2D) human joint detection from the image and (2) estimating a 3D pose from the 2D joints. H…
3D Human Pose Estimation3D Pose EstimationPose EstimationUnsupervised Single-shot Depth Estimation using Perceptual Reconstruction
Real-time estimation of actual object depth is an essential module for various autonomous system tasks such as 3D reconstruction, scene understanding and condition assessment. During the last decade of machine learning, …
3D ReconstructionDepth EstimationFace RecognitionScene UnderstandingUnsupervised 3D Human Pose Estimation via Conditional Multi-view Ancestral Sampling
We propose a method of estimating a 3D human pose from a single view without 3D supervision. The key to our method is to leverage the 2D diffusion priors of motion diffusion models (MDMs) pre-trained on large 2D human po…
Unsupervised 3D Human Pose Estimation3D Pose Estimation