Anchors Based Method for Fingertips Position Estimation from a Monocular RGB Image using Deep Neural Network
In Virtual, augmented, and mixed reality, the use of hand gestures is increasingly becoming popular to reduce the difference between the virtual and real world. The precise location of the fingertip is essential/crucial for a seamless experience. Much of the research work is based on using depth information for the estimation of the fingertips position. However, most of the work using RGB images for fingertips detection is limited to a single finger. The detection of multiple fingertips from a single RGB image is very challenging due to various factors. In this paper, we propose a deep neural network (DNN) based methodology to estimate the fingertips position. We christened this methodology as an Anchor based Fingertips Position Estimation (ABFPE), and it is a two-step process. The fingertips location is estimated using regression by computing the difference in the location of a fingertip from the nearest anchor point. The proposed framework performs the best with limited dependence on hand detection results. In our experiments on the SCUT-Ego-Gesture dataset, we achieved the fingertips detection error of 2.3552 pixels on a video frame with a resolution of $640 \times 480$ and about $92.98\%$ of test images have average pixel errors of five pixels.
Code (0)
등록된 구현이 없습니다.
Tasks
Hand DetectionMixed RealityPositionSimilar Papers 제목 키워드 기반
Learning Image-Adaptive Scale Fields for Metric Depth Recovery
Monocular depth estimation (MDE) typically produces depth estimations that are defined up to an unknown scale or shift. When only sparse metric anchors are available, recovering accurate metric depth becomes challenging …
Monocular Depth EstimationTwo-Stream Binocular Network: Accurate Near Field Finger Detection Based On Binocular Images
Fingertip detection plays an important role in human computer interaction. Previous works transform binocular images into depth images. Then depth-based hand pose estimation methods are used to predict 3D positions of fi…
Fingertip DetectionHand Pose EstimationPose EstimationReal-Time Tactile Grasp Force Sensing Using Fingernail Imaging via Deep Neural Networks
This paper has introduced a novel approach for the real-time estimation of 3D tactile forces exerted by human fingertips via vision only. The introduced approach is entirely monocular vision-based and does not require an…
Pose EstimationUnified Learning Approach for Egocentric Hand Gesture Recognition and Fingertip Detection
Head-mounted device-based human-computer interaction often requires egocentric recognition of hand gestures and fingertips detection. In this paper, a unified approach of egocentric hand gesture recognition and fingertip…
Fingertip DetectionGesture RecognitionHand Gesture RecognitionPositionDouble-Dot Network for Antipodal Grasp Detection
This paper proposes a new deep learning approach to antipodal grasp detection, named Double-Dot Network (DD-Net). It follows the recent anchor-free object detection framework, which does not depend on empirically pre-set…
object-detectionObject Detection