paper-with-me

홈 › Papers

Synergistic Global-space Camera and Human Reconstruction from Videos

2024-05-23 · CVPR 2024 1 · Yizhou Zhao, Tuanfeng Y. Wang, Bhiksha Raj, Min Xu, Jimei Yang, Chun-Hao Paul Huang

Remarkable strides have been made in reconstructing static scenes or human bodies from monocular videos. Yet, the two problems have largely been approached independently, without much synergy. Most visual SLAM methods can only reconstruct camera trajectories and scene structures up to scale, while most HMR methods reconstruct human meshes in metric scale but fall short in reasoning with cameras and scenes. This work introduces Synergistic Camera and Human Reconstruction (SynCHMR) to marry the best of both worlds. Specifically, we design Human-aware Metric SLAM to reconstruct metric-scale camera poses and scene point clouds using camera-frame HMR as a strong prior, addressing depth, scale, and dynamic ambiguities. Conditioning on the dense scene recovered, we further learn a Scene-aware SMPL Denoiser to enhance world-frame HMR by incorporating spatio-temporal coherency and dynamic scene constraints. Together, they lead to consistent reconstructions of camera trajectories, human meshes, and dense scene point clouds in a common world frame. Project page: https://paulchhuang.github.io/synchmr

📄 PDF Abstract BibTeX arXiv:2405.14855

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

TROPHIES: Temporal Reconstruction of Places, Humans, and Cameras from Multi-view Videos

2026-06-01 · Jinpeng Liu, Yukang Xu, Yutong Li, Xingyu Liu arxiv

Reconstructing humans and their surrounding environments in a globally consistent 4D space is essential for comprehensive perception. However, prior works typically assume single-view inputs or decouple humans, scenes, a…

Spatial Reasoning

Crowd3D++: Robust Monocular Crowd Reconstruction with Upright Space

2024-11-09 · Jing Huang, Hao Wen, Tianyi Zhou, Haozhe Lin 외

This paper aims to reconstruct hundreds of people's 3D poses, shapes, and locations from a single image with unknown camera parameters. Due to the small and highly varying 2D human scales, depth ambiguity, and perspectiv…

Reconstructing People, Places, and Cameras

2024-12-23 · CVPR 2025 1 · Lea Müller, Hongsuk Choi, Anthony Zhang, Brent Yi 외

We present "Humans and Structure from Motion" (HSfM), a method for jointly reconstructing multiple human meshes, scene point clouds, and camera parameters in a metric world coordinate system from a sparse set of uncalibr…

Camera Pose EstimationPose Estimation

DuoMo: Dual Motion Diffusion for World-Space Human Reconstruction

2026-03-03 · Yufu Wang, Evonne Ng, Soyong Shin, Rawal Khirodkar 외 arxiv

We present DuoMo, a generative method that recovers human motion in world-space coordinates from unconstrained videos with noisy or incomplete observations. Reconstructing such motion requires solving a fundamental trade…

WATCH: World-aware Allied Trajectory and pose reconstruction for Camera and Human

2025-09-04 · Qijun Ying, Zhongyuan Hu, Rui Zhang, Ronghui Li 외 arxiv

Global human motion reconstruction from in-the-wild monocular videos is increasingly demanded across VR, graphics, and robotics applications, yet requires accurate mapping of human poses from camera to world coordinates-…