Diffusion-based Pose Refinement and Muti-hypothesis Generation for 3D Human Pose Estimaiton
Previous probabilistic models for 3D Human Pose Estimation (3DHPE) aimed to enhance pose accuracy by generating multiple hypotheses. However, most of the hypotheses generated deviate substantially from the true pose. Compared to deterministic models, the excessive uncertainty in probabilistic models leads to weaker performance in single-hypothesis prediction. To address these two challenges, we propose a diffusion-based refinement framework called DRPose, which refines the output of deterministic models by reverse diffusion and achieves more suitable multi-hypothesis prediction for the current pose benchmark by multi-step refinement with multiple noises. To this end, we propose a Scalable Graph Convolution Transformer (SGCT) and a Pose Refinement Module (PRM) for denoising and refining. Extensive experiments on Human3.6M and MPI-INF-3DHP datasets demonstrate that our method achieves state-of-the-art performance on both single and multi-hypothesis 3DHPE. Code is available at https://github.com/KHB1698/DRPose.
Code (1)
Tasks
3D Human Pose EstimationDenoisingPose EstimationMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Fusing Urban Structure and Semantics: A Conditional Diffusion Model for Cross-City OD Matrix Generation
Accurate modeling of commuting flows is important for urban governance, traffic planning, and resource allocation. However, the combined influence of individual intentions, geographic constraints, and social dynamics lea…
From localized to well-mixed: How commuter interactions shape disease spread
Interactions between commuting individuals can lead to large-scale spreading of rumors, ideas, or disease, even though the commuters have no net displacement. The emergent dynamics depend crucially on the commuting distr…
Condition Errors Refinement in Autoregressive Image Generation with Diffusion Loss
Recent studies have explored autoregressive models for image generation, with promising results, and have combined diffusion models with autoregressive frameworks to optimize image generation via diffusion losses. In thi…
Image GenerationForward-Free Diffusion Language Models with BPTT-Free Looped Refinement
Diffusion language models generate text through iterative denoising, offering a powerful alternative to autoregressive generation. However, discrete language spaces lack a natural neighborhood structure for defining effe…
Iterative Token Evaluation and Refinement for Real-World Super-Resolution
Real-world image super-resolution (RWSR) is a long-standing problem as low-quality (LQ) images often have complex and unidentified degradations. Existing methods such as Generative Adversarial Networks (GANs) or continuo…
Image Super-ResolutionSuper-ResolutionTexture Synthesis