paper-with-me

홈 › Papers

HumanRAM: Feed-forward Human Reconstruction and Animation Model using Transformers

2025-06-03 · Zhiyuan Yu, Zhe Li, Hujun Bao, Can Yang, Xiaowei Zhou

3D human reconstruction and animation are long-standing topics in computer graphics and vision. However, existing methods typically rely on sophisticated dense-view capture and/or time-consuming per-subject optimization procedures. To address these limitations, we propose HumanRAM, a novel feed-forward approach for generalizable human reconstruction and animation from monocular or sparse human images. Our approach integrates human reconstruction and animation into a unified framework by introducing explicit pose conditions, parameterized by a shared SMPL-X neural texture, into transformer-based large reconstruction models (LRM). Given monocular or sparse input images with associated camera parameters and SMPL-X poses, our model employs scalable transformers and a DPT-based decoder to synthesize realistic human renderings under novel viewpoints and novel poses. By leveraging the explicit pose conditions, our model simultaneously enables high-quality human reconstruction and high-fidelity pose-controlled animation. Experiments show that HumanRAM significantly surpasses previous methods in terms of reconstruction accuracy, animation fidelity, and generalization performance on real-world datasets. Video results are available at https://zju3dv.github.io/humanram/.

📄 PDF Abstract BibTeX arXiv:2506.03118

Code (0)

등록된 구현이 없습니다.

Tasks

3D Human ReconstructionDecoder

Similar Papers 제목 키워드 기반

Real-Time Human Reconstruction and Animation using Feed-Forward Gaussian Splatting

2026-04-11 · Devdoot Chatterjee, Zakaria Laskar, C. V. Jawahar arxiv

We present HumanGS, a generalizable feed-forward Gaussian splatting framework for human 3D reconstruction and real-time animation from sparse multi-view RGB images and their associated SMPL-X poses. Unlike prior methods …

3D Reconstruction

SAT: Supervisor Regularization and Animation Augmentation for Two-process Monocular Texture 3D Human Reconstruction

2025-08-27 · Gangjian Zhang, Jian Shu, Nanjie Yao, Hao Wang arxiv

Monocular texture 3D human reconstruction aims to create a complete 3D digital avatar from just a single front-view human RGB image. However, the geometric ambiguity inherent in a single 2D image and the scarcity of 3D h…

3D Human Reconstruction

FRESA:Feedforward Reconstruction of Personalized Skinned Avatars from Few Images

2025-03-24 · Rong Wang, Fabian Prada, Ziyan Wang, Zhongshi Jiang 외

We present a novel method for reconstructing personalized 3D human avatars with realistic animation from only a few images. Due to the large variations in body shapes, poses, and cloth types, existing methods mostly requ…

3D CanonicalizationZero-shot Generalization

FRESA: Feedforward Reconstruction of Personalized Skinned Avatars from Few Images

2025-01-01 · CVPR 2025 1 · Rong Wang, Fabian Prada, Ziyan Wang, Zhongshi Jiang 외

We present a novel method for reconstructing personalized 3D human avatars with realistic animation from only a few images. Due to the large variations in body shapes, poses, and cloth types, existing methods mostly …

3D CanonicalizationZero-shot Generalization

AnimateAnyMesh++: A Flexible Feed-Forward Framework for High-Fidelity Text-Driven Mesh Animation

2026-04-29 · Zijie Wu, Chaohui Yu, Fan Wang, Xiang Bai arxiv

Recent advances in 4D content generation have attracted increasing attention, yet creating high-quality animated 3D models remains challenging due to the complexity of modeling spatio-temporal distributions and the scarc…