paper-with-me

홈 › Papers

SnapPose3D: Diffusion-Based Single-Frame 2D-to-3D Lifting of Human Poses

2026-04-29 · Alessandro Simoni, Riccardo Catalini, Davide Di Nucci, Guido Borghi, Davide Davoli, Lorenzo Garattoni, Gianpiero Francesca, Yuki Kawana, Roberto Vezzani arxiv

Depth ambiguity and joint uncertainty are the two main obstacles in obtaining accurate human pose predictions by 2D-to-3D lifting methods proposed in the literature. In particular, these issues are caused by 2D joint locations that can be mapped to multiple 3D positions, inducing multiple possible final poses. Following these considerations, we propose leveraging diffusion-based models generation capability to predict multiple hypotheses and aggregate them in a final accurate pose. Therefore, we introduce SnapPose3D, a pose-lifting framework trained deterministically to denoise 3D poses conditioned on both visual context and 2D pose features. SnapPose3D adopts a probabilistic approach during inference, generating multiple hypotheses through random sampling from a unit Gaussian distribution. Unlike most previous methods that address pose ambiguity by processing temporal sequences, SnapPose3D uses single frames as input, avoiding tracking and limiting computational cost, data acquisition complexity, and the need for online, real-time applications. We extensively evaluate SnapPose3D on well-known benchmarks for the 3D human pose estimation task showing its ability to generate and aggregate accurate hypotheses that lead to state-of-the-art results.

📄 PDF Abstract BibTeX arXiv:2604.26620

Code (0)

등록된 구현이 없습니다.

Tasks

3D Human Pose EstimationTemporal Sequences

Similar Papers 제목 키워드 기반

Unsupervised 3D Human Pose Estimation via Conditional Multi-view Ancestral Sampling

2026-05-15 · Ryohei Goto, Takuya Fujihashi, Shunsuke Saruwatari, Fumio Okura arxiv

We propose a method of estimating a 3D human pose from a single view without 3D supervision. The key to our method is to leverage the 2D diffusion priors of motion diffusion models (MDMs) pre-trained on large 2D human po…

Unsupervised 3D Human Pose Estimation3D Pose Estimation

NeuralLift-360: Lifting An In-the-wild 2D Photo to A 3D Object with 360° Views

2022-11-29 · Dejia Xu, Yifan Jiang, Peihao Wang, Zhiwen Fan 외

Virtual reality and augmented reality (XR) bring increasing demand for 3D content. However, creating high-quality 3D content requires tedious work that a human expert must do. In this work, we study the challenging task …

3D ReconstructionImage to 3DNeRFNovel View Synthesis+2

NeuralLift-360: Lifting an In-the-Wild 2D Photo to a 3D Object With 360deg Views

2023-01-01 · CVPR 2023 1 · Dejia Xu, Yifan Jiang, Peihao Wang, Zhiwen Fan 외

Virtual reality and augmented reality (XR) bring increasing demand for 3D content generation. However, creating high-quality 3D content requires tedious work from a human expert. In this work, we study the challengin…

DenoisingDepth EstimationNeRF

Online Action Recognition for Human Risk Prediction with Anticipated Haptic Alert via Wearables

2023-12-14 · Cheng Guo, Lorenzo Rapetti, Kourosh Darvish, Riccardo Grieco 외

This paper proposes a framework that combines online human state estimation, action recognition and motion prediction to enable early assessment and prevention of worker biomechanical risk during lifting tasks. The frame…

Action RecognitionMixture-of-Expertsmotion predictionState Estimation+1

PoseLifter: Absolute 3D human pose lifting network from a single noisy 2D human pose

2019-10-26 · Ju Yong Chang, Gyeongsik Moon, Kyoung Mu Lee

This study presents a new network (i.e., PoseLifter) that can lift a 2D human pose to an absolute 3D pose in a camera coordinate system. The proposed network estimates the absolute 3D location of a target subject and gen…

3D Human Pose EstimationPose Estimation