LG-Hand: Advancing 3D Hand Pose Estimation with Locally and Globally Kinematic Knowledge
3D hand pose estimation from RGB images suffers from the difficulty of obtaining the depth information. Therefore, a great deal of attention has been spent on estimating 3D hand pose from 2D hand joints. In this paper, we leverage the advantage of spatial-temporal Graph Convolutional Neural Networks and propose LG-Hand, a powerful method for 3D hand pose estimation. Our method incorporates both spatial and temporal dependencies into a single process. We argue that kinematic information plays an important role, contributing to the performance of 3D hand pose estimation. We thereby introduce two new objective functions, Angle and Direction loss, to take the hand structure into account. While Angle loss covers locally kinematic information, Direction loss handles globally kinematic one. Our LG-Hand achieves promising results on the First-Person Hand Action Benchmark (FPHAB) dataset. We also perform an ablation study to show the efficacy of the two proposed objective functions.
Code (0)
등록된 구현이 없습니다.
Tasks
3D Hand Pose EstimationHand Pose EstimationPose EstimationSimilar Papers 제목 키워드 기반
Self-Supervised 3D Hand Pose Estimation Through Training by Fitting
We present a self-supervision method for 3D hand pose estimation from depth maps. We begin with a neural network initialized with synthesized data and fine-tune it on real but unlabelled depth maps by minimizing a set of…
3D Hand Pose EstimationHand Pose EstimationPose EstimationAnyHand: A Large-Scale Synthetic Dataset for RGB(-D) Hand Pose Estimation
We present AnyHand, a large-scale synthetic dataset designed to advance the state of the art in 3D hand pose estimation. While recent works with foundation approaches have shown that scaling training data markedly improv…
3D Hand Pose EstimationMotion Estimation of Non-Holonomic Ground Vehicles From a Single Feature Correspondence Measured Over N Views
The planar motion of ground vehicles is often non-holonomic, which enables a solution of the two-view relative pose problem from a single point feature correspondence. Man-made environments such as underground parking lo…
Motion Estimationwgatools: an ultrafast toolkit for manipulating whole genome alignments
Summary: With the rapid development of long-read sequencing technologies, the era of individual complete genomes is approaching. We have developed wgatools, a cross-platform, ultrafast toolkit that supports a range of wh…
EgoPHI: Estimating Contact and Force from Egocentric Vision
Understanding hand-object interaction from egocentric vision is essential for modeling how people physically engage with the surrounding world. Yet reasoning about physically grounded interaction requires estimating the …