paper-with-me

홈 › Papers

MobilePoser: Real-Time Full-Body Pose Estimation and 3D Human Translation from IMUs in Mobile Consumer Devices

2025-04-16 · Vasco Xu, Chenfeng Gao, Henry Hoffmann, Karan Ahuja

There has been a continued trend towards minimizing instrumentation for full-body motion capture, going from specialized rooms and equipment, to arrays of worn sensors and recently sparse inertial pose capture methods. However, as these techniques migrate towards lower-fidelity IMUs on ubiquitous commodity devices, like phones, watches, and earbuds, challenges arise including compromised online performance, temporal consistency, and loss of global translation due to sensor noise and drift. Addressing these challenges, we introduce MobilePoser, a real-time system for full-body pose and global translation estimation using any available subset of IMUs already present in these consumer devices. MobilePoser employs a multi-stage deep neural network for kinematic pose estimation followed by a physics-based motion optimizer, achieving state-of-the-art accuracy while remaining lightweight. We conclude with a series of demonstrative applications to illustrate the unique potential of MobilePoser across a variety of fields, such as health and wellness, gaming, and indoor navigation to name a few.

📄 PDF Abstract BibTeX arXiv:2504.12492

Code (1)

SPICExLAB/MobilePoser 공식 구현 pytorch

Tasks

Pose EstimationTranslation

Similar Papers 제목 키워드 기반

BoDiffusion: Diffusing Sparse Observations for Full-Body Human Motion Synthesis

2023-04-21 · Angela Castillo, Maria Escobar, Guillaume Jeanneret, Albert Pumarola 외

Mixed reality applications require tracking the user's full-body motion to enable an immersive experience. However, typical head-mounted devices can only track head and hand movements, leading to a limited reconstruction…

Mixed RealityMotion Synthesis

AvatarReX: Real-time Expressive Full-body Avatars

2023-05-08 · Zerong Zheng, Xiaochen Zhao, Hongwen Zhang, Boning Liu 외

We present AvatarReX, a new method for learning NeRF-based full-body avatars from video data. The learnt avatar not only provides expressive control of the body, hands and the face together, but also supports real-time a…

DisentanglementNeRF

Learning Human-to-Humanoid Real-Time Whole-Body Teleoperation

2024-03-07 · Tairan He, Zhengyi Luo, Wenli Xiao, Chong Zhang 외

We present Human to Humanoid (H2O), a reinforcement learning (RL) based framework that enables real-time whole-body teleoperation of a full-sized humanoid robot with only an RGB camera. To create a large-scale retargeted…

Reinforcement Learning (RL)

Monocular Real-time Full Body Capture with Inter-part Correlations

2020-12-11 · CVPR 2021 1 · Yuxiao Zhou, Marc Habermann, Ikhsanul Habibie, Ayush Tewari 외

We present the first method for real-time full body capture that estimates shape and motion of body and hands together with a dynamic 3D face model from a single color image. Our approach uses a new neural network archit…

3D Hand Pose EstimationComputational EfficiencyFace Model

TaoAvatar: Real-Time Lifelike Full-Body Talking Avatars for Augmented Reality via 3D Gaussian Splatting

2025-03-21 · CVPR 2025 1 · Jianchuan Chen, Jingchuan Hu, Gaige Wang, Zhonghua Jiang 외

Realistic 3D full-body talking avatars hold great potential in AR, with applications ranging from e-commerce live streaming to holographic communication. Despite advances in 3D Gaussian Splatting (3DGS) for lifelike avat…

3DGS