Simultaneously Recovering Multi-Person Meshes and Multi-View Cameras with Human Semantics
Dynamic multi-person mesh recovery has broad applications in sports broadcasting, virtual reality, and video games. However, current multi-view frameworks rely on a time-consuming camera calibration procedure. In this work, we focus on multi-person motion capture with uncalibrated cameras, which mainly faces two challenges: one is that inter-person interactions and occlusions introduce inherent ambiguities for both camera calibration and motion capture; the other is that a lack of dense correspondences can be used to constrain sparse camera geometries in a dynamic multi-person scene. Our key idea is to incorporate motion prior knowledge to simultaneously estimate camera parameters and human meshes from noisy human semantics. We first utilize human information from 2D images to initialize intrinsic and extrinsic parameters. Thus, the approach does not rely on any other calibration tools or background features. Then, a pose-geometry consistency is introduced to associate the detected humans from different views. Finally, a latent motion prior is proposed to refine the camera parameters and human motions. Experimental results show that accurate camera parameters and human motions can be obtained through a one-step reconstruction. The code are publicly available at~\url{https://github.com/boycehbz/DMMR}.
Code (1)
Tasks
Camera CalibrationMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Multi-THuMBS: Multi-person Tracking of 3D Human Meshes Beyond Video Shots
Tracking multi-person 3D human meshes from in-the-wild videos is a highly challenging problem due to complex interactions, frequent occlusions, and severe truncation inherent in unconstrained environments. While recent a…
Camera Pose EstimationHuman Mesh RecoveryExploiting temporal context for 3D human pose estimation in the wild
We present a bundle-adjustment-based algorithm for recovering accurate 3D human pose and meshes from monocular videos. Unlike previous algorithms which operate on single frames, we show that reconstructing a person over …
3D Human Pose Estimation3D Pose EstimationMonocular 3D Human Pose EstimationPose EstimationMUG: Multi-human Graph Network for 3D Mesh Reconstruction from 2D Pose
Reconstructing multi-human body mesh from a single monocular image is an important but challenging computer vision problem. In addition to the individual body mesh models, we need to estimate relative 3D positions among …
3D Human Pose Estimation3D Multi-Person Human Pose Estimation3D Multi-Person Pose EstimationGraph Neural NetworkMonocular, One-stage, Regression of Multiple 3D People
This paper focuses on the regression of multiple 3D people from a single RGB image. Existing approaches predominantly follow a multi-stage pipeline that first detects people in bounding boxes and then independently regre…
3D Depth Estimation3D Human Pose Estimation3D Multi-Person Mesh Recovery3D Multi-Person Pose Estimation+2Body Meshes as Points
We consider the challenging multi-person 3D body mesh estimation task in this work. Existing methods are mostly two-stage based--one stage for person localization and the other stage for individual body mesh estimation, …
3D Human Pose Estimation3D Human Shape Estimation3D Multi-Person Pose Estimation3D Pose Estimation