Two-hand Global 3D Pose Estimation Using Monocular RGB
We tackle the challenging task of estimating global 3D joint locations for both hands via only monocular RGB input images. We propose a novel multi-stage convolutional neural network based pipeline that accurately segments and locates the hands despite occlusion between two hands and complex background noise and estimates the 2D and 3D canonical joint locations without any depth information. Global joint locations with respect to the camera origin are computed using the hand pose estimations and the actual length of the key bone with a novel projection algorithm. To train the CNNs for this new task, we introduce a large-scale synthetic 3D hand pose dataset. We demonstrate that our system outperforms previous works on 3D canonical hand pose estimation benchmark datasets with RGB-only information. Additionally, we present the first work that achieves accurate global 3D hand tracking on both hands using RGB-only inputs and provide extensive quantitative and qualitative evaluation.
Code (0)
등록된 구현이 없습니다.
Tasks
3D Canonical Hand Pose Estimation3D Hand Pose Estimation3D Pose EstimationHand Pose EstimationPose EstimationVocal Bursts Valence PredictionSimilar Papers 제목 키워드 기반
Dyn-HaMR: Recovering 4D Interacting Hand Motion from a Dynamic Camera
We propose Dyn-HaMR, to the best of our knowledge, the first approach to reconstruct 4D global hand motion from monocular videos recorded by dynamic cameras in the wild. Reconstructing accurate 3D hand meshes from monocu…
Simultaneous Localization and MappingGlobally Optimal Multi-Scale Monocular Hand-Eye Calibration Using Dual Quaternions
In this work, we present an approach for monocular hand-eye calibration from per-sensor ego-motion based on dual quaternions. Due to non-metrically scaled translations of monocular odometry, a scaling factor has to be es…
Translation${S}^{2}$Net: Accurate Panorama Depth Estimation on Spherical Surface
Monocular depth estimation is an ambiguous problem, thus global structural cues play an important role in current data-driven single-view depth estimation methods. Panorama images capture the complete spatial information…
DecoderDepth EstimationMonocular Depth EstimationPanoNormal: Monocular Indoor 360° Surface Normal Estimation
The presence of spherical distortion on the Equirectangular image is an acknowledged challenge in dense regression computer vision tasks, such as surface normal estimation. Recent advances in convolutional neural network…
Surface Normal EstimationScale-aware Insertion of Virtual Objects in Monocular Videos
In this paper, we propose a scale-aware method for inserting virtual objects with proper sizes into monocular videos. To tackle the scale ambiguity problem of geometry recovery from monocular videos, we estimate the glob…
Object