Beyond Weak Perspective for Monocular 3D Human Pose Estimation
We consider the task of 3D joints location and orientation prediction from a monocular video with the skinned multi-person linear (SMPL) model. We first infer 2D joints locations with an off-the-shelf pose estimation algorithm. We use the SPIN algorithm and estimate initial predictions of body pose, shape and camera parameters from a deep regression neural network. We then adhere to the SMPLify algorithm which receives those initial parameters, and optimizes them so that inferred 3D joints from the SMPL model would fit the 2D joints locations. This algorithm involves a projection step of 3D joints to the 2D image plane. The conventional approach is to follow weak perspective assumptions which use ad-hoc focal length. Through experimentation on the 3D Poses in the Wild (3DPW) dataset, we show that using full perspective projection, with the correct camera center and an approximated focal length, provides favorable results. Our algorithm has resulted in a winning entry for the 3DPW Challenge, reaching first place in joints orientation accuracy.
Code (0)
등록된 구현이 없습니다.
Tasks
3D Human Pose EstimationMonocular 3D Human Pose EstimationPose EstimationSimilar Papers 제목 키워드 기반
Error Bounds of Projection Models in Weakly Supervised 3D Human Pose Estimation
The current state-of-the-art in monocular 3D human pose estimation is heavily influenced by weakly supervised methods. These allow 2D labels to be used to learn effective 3D human pose recovery either directly from image…
3D Human Pose EstimationMonocular 3D Human Pose EstimationPose EstimationPosition+1An Empirical Study of Monocular Human Body Measurement Under Weak Calibration
Estimating human body measurements from monocular RGB imagery remains challenging due to scale ambiguity, viewpoint sensitivity, and the absence of explicit depth information. This work presents a systematic empirical st…
EPOCH: Jointly Estimating the 3D Pose of Cameras and Humans
Monocular Human Pose Estimation (HPE) aims at determining the 3D positions of human joints from a single 2D image captured by a camera. However, a single 2D point in the image may correspond to multiple points in 3D spac…
3D Pose EstimationPose EstimationDeep Physics-aware Inference of Cloth Deformation for Monocular Human Performance Capture
Recent monocular human performance capture approaches have shown compelling dense tracking results of the full body from a single RGB camera. However, existing methods either do not estimate clothing at all or model clot…
Synthetic Training for Monocular Human Mesh Recovery
Recovering 3D human mesh from monocular images is a popular topic in computer vision and has a wide range of applications. This paper aims to estimate 3D mesh of multiple body parts (e.g., body, hands) with large-scale d…
Computational EfficiencyHuman Mesh Recovery