paper-with-me

홈 › Papers

Neural Voting Field for Camera-Space 3D Hand Pose Estimation

2023-05-07 · CVPR 2023 1 · Lin Huang, Chung-Ching Lin, Kevin Lin, Lin Liang, Lijuan Wang, Junsong Yuan, Zicheng Liu

We present a unified framework for camera-space 3D hand pose estimation from a single RGB image based on 3D implicit representation. As opposed to recent works, most of which first adopt holistic or pixel-level dense regression to obtain relative 3D hand pose and then follow with complex second-stage operations for 3D global root or scale recovery, we propose a novel unified 3D dense regression scheme to estimate camera-space 3D hand pose via dense 3D point-wise voting in camera frustum. Through direct dense modeling in 3D domain inspired by Pixel-aligned Implicit Functions for 3D detailed reconstruction, our proposed Neural Voting Field (NVF) fully models 3D dense local evidence and hand global geometry, helping to alleviate common 2D-to-3D ambiguities. Specifically, for a 3D query point in camera frustum and its pixel-aligned image feature, NVF, represented by a Multi-Layer Perceptron, regresses: (i) its signed distance to the hand surface; (ii) a set of 4D offset vectors (1D voting weight and 3D directional vector to each hand joint). Following a vote-casting scheme, 4D offset vectors from near-surface points are selected to calculate the 3D hand joint coordinates by a weighted average. Experiments demonstrate that NVF outperforms existing state-of-the-art algorithms on FreiHAND dataset for camera-space 3D hand pose estimation. We also adapt NVF to the classic task of root-relative 3D hand pose estimation, for which NVF also obtains state-of-the-art results on HO3D dataset.

📄 PDF Abstract BibTeX arXiv:2305.04328

Code (0)

등록된 구현이 없습니다.

Tasks

3D Hand Pose EstimationHand Pose EstimationPose Estimationregression

Similar Papers 제목 키워드 기반

Sequential Voting with Relational Box Fields for Active Object Detection

2021-10-21 · CVPR 2022 1 · Qichen Fu, Xingyu Liu, Kris M. Kitani

A key component of understanding hand-object interactions is the ability to identify the active object -- the object that is being manipulated by the human hand. In order to accurately localize the active object, any met…

Active Object DetectionDecision MakingImitation LearningObject+3

Where to drive: free space detection with one fisheye camera

2020-11-11 · Tobias Scheck, Adarsh Mallandur, Christian Wiede, Gangolf Hirtz

The development in the field of autonomous driving goes hand in hand with ever new developments in the field of image processing and machine learning methods. In order to fully exploit the advantages of deep learning, it…

Autonomous DrivingDeep Learning

Exploiting Vector Fields for Geometric Rectification of Distorted Document Images

2018-09-01 · ECCV 2018 9 · Gaofeng MENG, Yuanqi SU, Ying Wu, Shiming Xiang 외

This paper proposes a segment-free method for geometric rectification of a distorted document image captured by a hand-held camera. The method can recover the 3D page shape by exploiting the intrinsic vector fields of th…

Camera Pose Voting for Large-Scale Image-Based Localization

2015-12-01 · ICCV 2015 12 · Bernhard Zeisl, Torsten Sattler, Marc Pollefeys

Image-based localization approaches aim to determine the camera pose from which an image was taken. Finding correct 2D-3D correspondences between query image features and 3D points in the scene model becomes harder as th…

Camera Pose EstimationImage-Based LocalizationPose Estimation

On Linear Structure From Motion for Light Field Cameras

2015-12-01 · ICCV 2015 12 · Ole Johannsen, Antonin Sulc, Bastian Goldluecke

We present a novel approach to relative pose estimation which is tailored to 4D light field cameras. From the relationships between scene geometry and light field structure and an analysis of the light field projection i…

Point cloud reconstructionPose Estimation