paper-with-me

홈 › Papers

Neural Monocular 3D Human Motion Capture with Physical Awareness

2021-05-03 · Soshi Shimada, Vladislav Golyanik, Weipeng Xu, Patrick Pérez, Christian Theobalt

We present a new trainable system for physically plausible markerless 3D human motion capture, which achieves state-of-the-art results in a broad range of challenging scenarios. Unlike most neural methods for human motion capture, our approach, which we dub physionical, is aware of physical and environmental constraints. It combines in a fully differentiable way several key innovations, i.e., 1. a proportional-derivative controller, with gains predicted by a neural network, that reduces delays even in the presence of fast motions, 2. an explicit rigid body dynamics model and 3. a novel optimisation layer that prevents physically implausible foot-floor penetration as a hard constraint. The inputs to our system are 2D joint keypoints, which are canonicalised in a novel way so as to reduce the dependency on intrinsic camera parameters -- both at train and test time. This enables more accurate global translation estimation without generalisability loss. Our model can be finetuned only with 2D annotations when the 3D annotations are not available. It produces smooth and physically principled 3D motions in an interactive frame rate in a wide variety of challenging scenes, including newly recorded ones. Its advantages are especially noticeable on in-the-wild sequences that significantly differ from common 3D pose estimation benchmarks such as Human 3.6M and MPI-INF-3DHP. Qualitative results are available at http://gvv.mpi-inf.mpg.de/projects/PhysAware/

📄 PDF Abstract BibTeX arXiv:2105.01057

Code (0)

등록된 구현이 없습니다.

Tasks

3D Pose EstimationPose Estimation

Methods 이 논문이 사용한 방법론

AWARE We propose to theoretically and empirically examine the effect of incorporating weighting schemes into walk-aggregating GNNs. To this end, we propose a simple, interpretable, and…

Similar Papers 제목 키워드 기반

Gravity-Aware Monocular 3D Human-Object Reconstruction

2021-08-19 · ICCV 2021 10 · Rishabh Dabral, Soshi Shimada, Arjun Jain, Christian Theobalt 외

This paper proposes GraviCap, i.e., a new approach for joint markerless 3D human motion capture and object trajectory estimation from monocular RGB videos. We focus on scenes with objects partially observed during a free…

Human-Object Interaction DetectionObjectObject Reconstruction

ProxyCap: Real-time Monocular Full-body Capture in World Space via Human-Centric Proxy-to-Motion Learning

2023-07-03 · CVPR 2024 1 · Yuxiang Zhang, Hongwen Zhang, Liangxiao Hu, Jiajun Zhang 외

Learning-based approaches to monocular motion capture have recently shown promising results by learning to regress in a data-driven manner. However, due to the challenges in data collection and network designs, it remain…

3D Human Pose Estimation

PhysPT: Physics-aware Pretrained Transformer for Estimating Human Dynamics from Monocular Videos

2024-04-05 · CVPR 2024 1 · Yufei Zhang, Jeffrey O. Kephart, Zijun Cui, Qiang Ji

While current methods have shown promising progress on estimating 3D human motion from monocular videos, their motion estimates are often physically unrealistic because they mainly consider kinematics. In this paper, we …

Action RecognitionDecoderHuman DynamicsTemporal Action Localization

Neural MoCon: Neural Motion Control for Physically Plausible Human Motion Capture

2022-03-26 · CVPR 2022 1 · Buzhen Huang, Liang Pan, Yuan Yang, Jingyi Ju 외

Due to the visual ambiguity, purely kinematic formulations on monocular human motion capture are often physically incorrect, biomechanically implausible, and can not reconstruct accurate interactions. In this work, we fo…

Decoder

A Plug-and-Play Physical Motion Restoration Approach for In-the-Wild High-Difficulty Motions

2024-12-23 · Youliang Zhang, Ronghui Li, Yachao Zhang, Liang Pan 외

Extracting physically plausible 3D human motion from videos is a critical task. Although existing simulation-based motion imitation methods can enhance the physical quality of daily motions estimated from monocular video…