paper-with-me

홈 › Papers

Enhancing Monocular 3D Hand Reconstruction with Learned Texture Priors

2025-08-13 · Giorgos Karvounas, Nikolaos Kyriazis, Iason Oikonomidis, Georgios Pavlakos, Antonis A. Argyros arxiv

We revisit the role of texture in monocular 3D hand reconstruction, not as an afterthought for photorealism, but as a dense, spatially grounded cue that can actively support pose and shape estimation. Our observation is simple: even in high-performing models, the overlay between predicted hand geometry and image appearance is often imperfect, suggesting that texture alignment may be an underused supervisory signal. We propose a lightweight texture module that embeds per-pixel observations into UV texture space and enables a novel dense alignment loss between predicted and observed hand appearances. Our approach assumes access to a differentiable rendering pipeline and a model that maps images to 3D hand meshes with known topology, allowing us to back-project a textured hand onto the image and perform pixel-based alignment. The module is self-contained and easily pluggable into existing reconstruction pipelines. To isolate and highlight the value of texture-guided supervision, we augment HaMeR, a high-performing yet unadorned transformer architecture for 3D hand pose estimation. The resulting system improves both accuracy and realism, demonstrating the value of appearance-guided alignment in hand reconstruction.

📄 PDF Abstract BibTeX arXiv:2508.09629

Code (0)

등록된 구현이 없습니다.

Tasks

3D Hand Pose Estimation

Similar Papers 제목 키워드 기반

HiFiHR: Enhancing 3D Hand Reconstruction from a Single Image via High-Fidelity Texture

2023-08-25 · Jiayin Zhu, Zhuoran Zhao, Linlin Yang, Angela Yao

We present HiFiHR, a high-fidelity hand reconstruction approach that utilizes render-and-compare in the learning-based framework from a single image, capable of generating visually plausible and accurate 3D hand meshes w…

Learning Deeply Supervised Good Features to Match for Dense Monocular Reconstruction

2017-11-16 · Chamara Saroj Weerasekera, Ravi Garg, Yasir Latif, Ian Reid

Visual SLAM (Simultaneous Localization and Mapping) methods typically rely on handcrafted visual features or raw RGB values for establishing correspondences between images. These features, while suitable for sparse mappi…

Depth EstimationMonocular ReconstructionSimultaneous Localization and Mapping

TexHOI: Reconstructing Textures of 3D Unknown Objects in Monocular Hand-Object Interaction Scenes

2025-01-07 · Alakh Aggarwal, Ningna Wang, Xiaohu Guo

Reconstructing 3D models of dynamic, real-world objects with high-fidelity textures from monocular frame sequences has been a challenging problem in recent years. This difficulty stems from factors such as shadows, indir…

Object

OAHuman: Occlusion-Aware 3D Human Reconstruction from Monocular Images

2026-03-15 · Yuanwang Yang, Hongliang Liu, Muxin Zhang, Nan Ma 외 arxiv

Monocular 3D human reconstruction in real-world scenarios remains highly challenging due to frequent occlusions from surrounding objects, people, or image truncation. Such occlusions lead to missing geometry and unreliab…

3D Human Reconstruction

Consistent 3D Hand Reconstruction in Video via self-supervised Learning

2022-01-24 · Zhigang Tu, Zhisheng Huang, Yujin Chen, Di Kang 외

We present a method for reconstructing accurate and consistent 3D hands from a monocular video. We observe that detected 2D hand keypoints and the image texture provide important cues about the geometry and texture of th…

Self-Supervised Learning