paper-with-me

홈 › Papers

An End-to-end Framework for Unconstrained Monocular 3D Hand Pose Estimation

2019-11-28 · Sanjeev Sharma, Shaoli Huang, DaCheng Tao

This work addresses the challenging problem of unconstrained 3D hand pose estimation using monocular RGB images. Most of the existing approaches assume some prior knowledge of hand (such as hand locations and side information) is available for 3D hand pose estimation. This restricts their use in unconstrained environments. We, therefore, present an end-to-end framework that robustly predicts hand prior information and accurately infers 3D hand pose by learning ConvNet models while only using keypoint annotations. To achieve robustness, the proposed framework uses a novel keypoint-based method to simultaneously predict hand regions and side labels, unlike existing methods that suffer from background color confusion caused by using segmentation or detection-based technology. Moreover, inspired by the biological structure of the human hand, we introduce two geometric constraints directly into the 3D coordinates prediction that further improves its performance in a weakly-supervised training. Experimental results show that our proposed framework not only performs robustly on unconstrained setting, but also outperforms the state-of-art methods on standard benchmark datasets.

📄 PDF Abstract BibTeX arXiv:1911.12501

Code (0)

등록된 구현이 없습니다.

Tasks

3D Hand Pose EstimationHand Pose EstimationPose Estimation

Similar Papers 제목 키워드 기반

Region Deformer Networks for Unsupervised Depth Estimation from Unconstrained Monocular Videos

2019-02-26 · Haofei Xu, Jianmin Zheng, Jianfei Cai, Juyong Zhang

While learning based depth estimation from images/videos has achieved substantial progress, there still exist intrinsic limitations. Supervised methods are limited by a small amount of ground truth or labeled data and un…

Depth Estimation

HandsOnWorld: Unconstrained Egocentric Video Generation with Camera-Disentangled Hand Control

2026-07-02 · Yushuo Chen, Xiaoyu Shi, Xiaoshi Wu, Xintao Wang 외 arxiv

We present HandsOnWorld, a framework for hand-controlled egocentric video generation that learns directly from unconstrained monocular video. Prior generators depend on 3D hand annotations from multi-view or marker-based…

Video Generation

Unconstrained Monocular 3D Human Pose Estimation by Action Detection and Cross-Modality Regression Forest

2013-06-01 · CVPR 2013 6 · Tsz-Ho Yu, Tae-Kyun Kim, Roberto Cipolla

This work addresses the challenging problem of unconstrained 3D human pose estimation (HPE) from a novel perspective. Existing approaches struggle to operate in realistic applications, mainly due to their scene-dependent…

2D Pose Estimation3D Human Pose EstimationAction DetectionMonocular 3D Human Pose Estimation+3

Towards unconstrained joint hand-object reconstruction from RGB videos

2021-08-16 · Yana Hasson, Gül Varol, Ivan Laptev, Cordelia Schmid

Our work aims to obtain 3D reconstruction of hands and manipulated objects from monocular videos. Reconstructing hand-object manipulations holds a great potential for robotics and learning from human demonstrations. The …

3D Hand Pose Estimation3D Reconstructionhand-object poseHand Pose Estimation+7

Toward a Real-Time Framework for Accurate Monocular 3D Human Pose Estimation with Geometric Priors

2025-07-21 · Mohamed Adjel arxiv

Monocular 3D human pose estimation remains a challenging and ill-posed problem, particularly in real-time settings and unconstrained environments. While direct imageto-3D approaches require large annotated datasets and h…

Monocular 3D Human Pose EstimationKeypoint Detection3D Pose Estimation