paper-with-me

홈 › Papers

FactorizedHMR: A Hybrid Framework for Video Human Mesh Recovery

2026-05-14 · Patrick Kwon, Chen Chen arxiv

Human Mesh Recovery (HMR) is fundamentally ambiguous: under occlusion or weak depth cues, multiple 3D bodies can explain the same image evidence. This ambiguity is not uniform across the body, as torso pose and root structure are often relatively well constrained, whereas distal articulations such as the arms and legs are more uncertain. Building on this observation, we propose FactorizedHMR, a two-stage framework that treats these two regimes differently. A deterministic regression module first recovers a stable torso-root anchor, and a probabilistic flow-matching module then completes the remaining non-torso articulation. To make this completion reliable, we combine a composite target representation with geometry-aware supervision and feature-aware classifier-free guidance, preserving the torso-root anchor while improving single-reference recovery of ambiguity-prone articulation. We also introduce a synthetic data pipeline that provides the paired image-camera-motion supervision under diverse viewpoints. Across camera-space and world-space benchmarks, FactorizedHMR remains competitive with strong baselines, with the clearest gains in occlusion-heavy recovery and drift-sensitive world-space metrics.

📄 PDF Abstract BibTeX arXiv:2605.14854

Code (0)

등록된 구현이 없습니다.

Tasks

Human Mesh Recovery

Similar Papers 제목 키워드 기반

SplattingAvatar: Realistic Real-Time Human Avatars with Mesh-Embedded Gaussian Splatting

2024-03-08 · CVPR 2024 1 · Zhijing Shao, Zhaolong Wang, Zhuang Li, Duotun Wang 외

We present SplattingAvatar, a hybrid 3D representation of photorealistic human avatars with Gaussian Splatting embedded on a triangle mesh, which renders over 300 FPS on a modern GPU and 30 FPS on a mobile device. We dis…

GPU

GoMAvatar: Efficient Animatable Human Modeling from Monocular Video Using Gaussians-on-Mesh

2024-04-11 · CVPR 2024 1 · Jing Wen, Xiaoming Zhao, Zhongzheng Ren, Alexander G. Schwing 외

We introduce GoMAvatar, a novel approach for real-time, memory-efficient, high-quality animatable human modeling. GoMAvatar takes as input a single monocular video to create a digital avatar capable of re-articulation in…

Computational Efficiency

Expressive Whole-Body 3D Gaussian Avatar

2024-07-31 · Gyeongsik Moon, Takaaki Shiratori, Shunsuke Saito

Facial expression and hand motions are necessary to express our emotions and interact with the world. Nevertheless, most of the 3D human avatars modeled from a casually captured video only support body motions without fa…

3DGSDiversity

DiffMesh: A Motion-aware Diffusion Framework for Human Mesh Recovery from Videos

2023-03-23 · Ce Zheng, Xianpeng Liu, Qucheng Peng, Tianfu Wu 외

Human mesh recovery (HMR) provides rich human body information for various real-world applications. While image-based HMR methods have achieved impressive results, they often struggle to recover humans in dynamic scenari…

3D Human Pose EstimationHuman Mesh Recovery

PC-HMR: Pose Calibration for 3D Human Mesh Recovery from 2D Images/Videos

2021-03-16 · Tianyu Luan, Yali Wang, Junhao Zhang, Zhe Wang 외

The end-to-end Human Mesh Recovery (HMR) approach has been successfully used for 3D body reconstruction. However, most HMR-based frameworks reconstruct human body by directly learning mesh parameters from images or video…

3D Human Pose EstimationHuman Mesh Recovery