Lightweight Multi-person Total Motion Capture Using Sparse Multi-view Cameras
Multi-person total motion capture is extremely challenging when it comes to handle severe occlusions, different reconstruction granularities from body to face and hands, drastically changing observation scales and fast body movements. To overcome these challenges above, we contribute a lightweight total motion capture system for multi-person interactive scenarios using only sparse multi-view cameras. By contributing a novel hand and face bootstrapping algorithm, our method is capable of efficient localization and accurate association of the hands and faces even on severe occluded occasions. We leverage both pose regression and keypoints detection methods and further propose a unified two-stage parametric fitting method for achieving pixel-aligned accuracy. Moreover, for extremely self-occluded poses and close interactions, a novel feedback mechanism is proposed to propagate the pixel-aligned reconstructions into the next frame for more accurate association. Overall, we propose the first light-weight total capture system and achieves fast, robust and accurate multi-person total motion capture performance. The results and experiments show that our method achieves more accurate results than existing methods under sparse-view setups.
Code (0)
등록된 구현이 없습니다.
Tasks
3D Multi-Person Pose EstimationSimilar Papers 제목 키워드 기반
Monocular Total Capture: Posing Face, Body, and Hands in the Wild
We present the first method to capture the 3D total motion of a target person from a monocular view input. Given an image or a monocular video, our method reconstructs the motion from body, face, and fingers represented …
3D Human Pose EstimationHand Pose EstimationMonocular 3D Human Pose EstimationA Multimodal Motion-Captured Corpus of Matched and Mismatched Extravert-Introvert Conversational Pairs
This paper presents a new corpus, the Personality Dyads Corpus, consisting of multimodal data for three conversations between three personality-matched, two-person dyads (a total of 9 separate dialogues). Participants we…
Ultra Inertial Poser: Scalable Motion Capture and Tracking from Sparse Inertial Sensors and Ultra-Wideband Ranging
While camera-based capture systems remain the gold standard for recording human motion, learning-based tracking systems based on sparse wearable sensors are gaining popularity. Most commonly, they use inertial sensors, w…
Pose EstimationPersonalizing Causal Audio-Driven Facial Motion via Dynamic Multi-modal Retrieval
Audio-driven facial animation is essential for immersive digital interaction, yet existing frameworks fail to reconcile real-time streaming with high-fidelity personalization. Current methods often rely on latency-induci…
WorldPose: A World Cup Dataset for Global 3D Human Pose Estimation
We present WorldPose, a novel dataset for advancing research in multi-person global pose estimation in the wild, featuring footage from the 2022 FIFA World Cup. While previous datasets have primarily focused on local pos…
3D Human Pose EstimationGlobal 3D Human Pose EstimationPose Estimation