paper-with-me

홈 › Papers

DexAvatar: 3D Sign Language Reconstruction with Hand and Body Pose Priors

2025-12-24 · Kaustubh Kundu, Hrishav Bakul Barua, Lucy Robertson-Bell, Zhixi Cai, Kalin Stefanov arxiv

The trend in sign language generation is centered around data-driven generative methods that require vast amounts of precise 2D and 3D human pose data to achieve an acceptable generation quality. However, currently, most sign language datasets are video-based and limited to automatically reconstructed 2D human poses (i.e., keypoints) and lack accurate 3D information. Furthermore, existing state-of-the-art for automatic 3D human pose estimation from sign language videos is prone to self-occlusion, noise, and motion blur effects, resulting in poor reconstruction quality. In response to this, we introduce DexAvatar, a novel framework to reconstruct bio-mechanically accurate fine-grained hand articulations and body movements from in-the-wild monocular sign language videos, guided by learned 3D hand and body priors. DexAvatar achieves strong performance in the SGNify motion capture dataset, the only benchmark available for this task, reaching an improvement of 35.11% in the estimation of body and hand poses compared to the state-of-the-art. The official website of this work is: https://github.com/kaustesseract/DexAvatar.

📄 PDF Abstract BibTeX arXiv:2512.21054

Code (0)

등록된 구현이 없습니다.

Tasks

3D Human Pose Estimation

Similar Papers 제목 키워드 기반

Independent Sign Language Recognition with 3D Body, Hands, and Face Reconstruction

2020-11-24 · Agelos Kratimenos, Georgios Pavlakos, Petros Maragos

Independent Sign Language Recognition is a complex visual recognition problem that combines several challenging tasks of Computer Vision due to the necessity to exploit and fuse information from hand gestures, body featu…

3D Action Recognition3D ReconstructionAction RecognitionFace Reconstruction+2

Tamaththul3D: High-Fidelity 3D Saudi Sign Language Avatars from Monocular Video

2026-05-06 · Eyad Alghamdi, Sattam Altuuaim, Obay Ghulam, Abdulrahman Qutah 외 arxiv

Existing 3D sign language avatar reconstruction methods are developed and evaluated exclusively on Western sign languages, and no 3D parametric annotations exist for any Arabic Sign Language dataset, a gap that blocks th…

Stratified Avatar Generation from Sparse Observations

2024-05-30 · CVPR 2024 1 · Han Feng, Wenchao Ma, Quankai Gao, Xianwei Zheng 외

Estimating 3D full-body avatars from AR/VR devices is essential for creating immersive experiences in AR/VR applications. This task is challenging due to the limited input from Head Mounted Devices, which capture only sp…

Decoder

DanceHMR: Hand-Aware Whole-Body Human Mesh Recovery from Monocular Videos

2026-05-18 · Wenhao Shen, Ming Zhou, Hengyuan Zhang, Siyuan Bian 외 arxiv

Monocular video human mesh recovery is essential for digital humans, avatar animation, and embodied simulation, where both temporal stability and expressive whole-body motion are required. Existing video HMR methods prod…

Human Mesh Recovery

BoDiffusion: Diffusing Sparse Observations for Full-Body Human Motion Synthesis

2023-04-21 · Angela Castillo, Maria Escobar, Guillaume Jeanneret, Albert Pumarola 외

Mixed reality applications require tracking the user's full-body motion to enable an immersive experience. However, typical head-mounted devices can only track head and hand movements, leading to a limited reconstruction…

Mixed RealityMotion Synthesis