paper-with-me

홈 › Papers

Semantic Belief-State World Model for 3D Human Motion Prediction

2026-01-07 · Sarim Chaudhry arxiv

Human motion prediction has traditionally been framed as a sequence regression problem where models extrapolate future joint coordinates from observed pose histories. While effective over short horizons this approach does not separate observation reconstruction with dynamics modeling and offers no explicit representation of the latent causes governing motion. As a result, existing methods exhibit compounding drift, mean-pose collapse, and poorly calibrated uncertainty when rolled forward beyond the training regime. Here we propose a Semantic Belief-State World Model (SBWM) that reframes human motion prediction as latent dynamical simulation on the human body manifold. Rather than predicting poses directly, SBWM maintains a recurrent probabilistic belief state whose evolution is learned independently of pose reconstruction and explicitly aligned with the SMPL-X anatomical parameterization. This alignment imposes a structural information bottleneck that prevents the latent state from encoding static geometry or sensor noise, forcing it to capture motion dynamics, intent, and control-relevant structure. Inspired by belief-state world models developed for model-based reinforcement learning, SBWM adapts stochastic latent transitions and rollout-centric training to the domain of human motion. In contrast to RSSM-based, transformer, and diffusion approaches optimized for reconstruction fidelity, SBWM prioritizes stable forward simulation. We demonstrate coherent long-horizon rollouts, and competitive accuracy at substantially lower computational cost. These results suggest that treating the human body as part of the world models state space rather than its output fundamentally changes how motion is simulated, and predicted.

📄 PDF Abstract BibTeX arXiv:2601.03517

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

Inferring World Belief States in Dynamic Real-World Environments

2026-04-13 · Jack Kolb, Aditya Garg, Nikolai Warner, Karen M. Feigh arxiv

We investigate estimating a human's world belief state using a robot's observations in a dynamic, 3D, and partially observable environment. The methods are grounded in mental model theory, which posits that human decisio…

Decision Making

Grounding Language about Belief in a Bayesian Theory-of-Mind

2024-02-16 · Lance Ying, Tan Zhi-Xuan, Lionel Wong, Vikash Mansinghka 외

Despite the fact that beliefs are mental states that cannot be directly observed, humans talk about each others' beliefs on a regular basis, often using rich compositional language to describe what others think and know.…

Attribute

AffectVerse: Emotional World Models for Multimodal Affective Computing

2026-05-19 · Bo Zhao, Fanghua Ye, Yixin Ji, Sicheng Zhao 외 arxiv

Humans infer emotions by integrating observed multimodal cues with expectations about how affective states may unfold. Existing multimodal large language models (MLLMs), however, often treat emotion recognition as static…

Emotion Recognition

Task-assisted Motion Planning in Partially Observable Domains

2019-08-27 · Antony Thomas, Sunny Amatya, Fulvio Mastrogiovanni, Marco Baglietto

We present an integrated Task-Motion Planning framework for robot navigation in belief space. Autonomous robots operating in real world complex scenarios require planning in the discrete (task) space and the continuous (…

Motion PlanningRobot Navigation

OmniToM: Benchmarking Theory of Mind in LLMs via Explicit Belief Modeling

2026-05-25 · Adam Bawatneh, Sagar Sapkota, Amrit Singh Bedi, Santu Karmaker 외 arxiv

Theory of Mind (ToM), the ability to infer others' knowledge, intentions, and emotions, is commonly evaluated in large language models (LLMs) using end-point question answering, where performance is judged solely by the …

Question Answering