paper-with-me

홈 › Papers

DPoser-X: Diffusion Model as Robust 3D Whole-body Human Pose Prior

2025-08-01 · Junzhe Lu, Jing Lin, Hongkun Dou, Ailing Zeng, Yue Deng, Xian Liu, Zhongang Cai, Lei Yang, Yulun Zhang, Haoqian Wang, Ziwei Liu arxiv

We present DPoser-X, a diffusion-based prior model for 3D whole-body human poses. Building a versatile and robust full-body human pose prior remains challenging due to the inherent complexity of articulated human poses and the scarcity of high-quality whole-body pose datasets. To address these limitations, we introduce a Diffusion model as body Pose prior (DPoser) and extend it to DPoser-X for expressive whole-body human pose modeling. Our approach unifies various pose-centric tasks as inverse problems, solving them through variational diffusion sampling. To enhance performance on downstream applications, we introduce a novel truncated timestep scheduling method specifically designed for pose data characteristics. We also propose a masked training mechanism that effectively combines whole-body and part-specific datasets, enabling our model to capture interdependencies between body parts while avoiding overfitting to specific actions. Extensive experiments demonstrate DPoser-X's robustness and versatility across multiple benchmarks for body, hand, face, and full-body pose modeling. Our model consistently outperforms state-of-the-art alternatives, establishing a new benchmark for whole-body human pose prior modeling.

📄 PDF Abstract BibTeX arXiv:2508.00599

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

DPoser: Diffusion Model as Robust 3D Human Pose Prior

2023-12-09 · Junzhe Lu, Jing Lin, Hongkun Dou, Ailing Zeng 외

This work targets to construct a robust human pose prior. However, it remains a persistent challenge due to biomechanical constraints and diverse human movements. Traditional priors like VAEs and NDFs often exhibit short…

DenoisingHuman Mesh RecoveryScheduling

Egocentric Whole-Body Motion Capture with FisheyeViT and Diffusion-Based Motion Refinement

2023-11-28 · CVPR 2024 1 · Jian Wang, Zhe Cao, Diogo Luvizon, Lingjie Liu 외

In this work, we explore egocentric whole-body motion capture using a single fisheye camera, which simultaneously estimates human body and hand motion. This task presents significant challenges due to three factors: the …

Egocentric Pose EstimationHand DetectionHand Pose EstimationPose Estimation+1

GraspDiffusion: Synthesizing Realistic Whole-body Hand-Object Interaction

2024-10-17 · Patrick Kwon, Hanbyul Joo

Recent generative models can synthesize high-quality images but often fail to generate humans interacting with objects using their hands. This arises mostly from the model's misunderstanding of such interactions, and the…

Human-Object Interaction DetectionImage GenerationObject

DreamControl: Human-Inspired Whole-Body Humanoid Control for Scene Interaction via Guided Diffusion

2025-09-17 · Dvij Kalaria, Sudarshan S Harithas, Pushkal Katara, Sangkyung Kwak 외 arxiv

We introduce DreamControl, a novel methodology for learning autonomous whole-body humanoid skills. DreamControl leverages the strengths of diffusion models and Reinforcement Learning (RL): our core innovation is the use …

Reinforcement Learning

REWIND: Real-Time Egocentric Whole-Body Motion Diffusion with Exemplar-Based Identity Conditioning

2025-04-07 · CVPR 2025 1 · Jihyun Lee, Weipeng Xu, Alexander Richard, Shih-En Wei 외

We present REWIND (Real-Time Egocentric Whole-Body Motion Diffusion), a one-step diffusion model for real-time, high-fidelity human motion estimation from egocentric image inputs. While an existing method for egocentric …

DenoisingMotion Estimation