paper-with-me

홈 › Papers

How Much Do Large Language Models Know about Human Motion? A Case Study in 3D Avatar Control

2025-05-23 · Kunhang Li, Jason Naradowsky, Yansong Feng, Yusuke Miyao

We explore Large Language Models (LLMs)' human motion knowledge through 3D avatar control. Given a motion instruction, we prompt LLMs to first generate a high-level movement plan with consecutive steps (High-level Planning), then specify body part positions in each step (Low-level Planning), which we linearly interpolate into avatar animations as a clear verification lens for human evaluators. Through carefully designed 20 representative motion instructions with full coverage of basic movement primitives and balanced body part usage, we conduct comprehensive evaluations including human assessment of both generated animations and high-level movement plans, as well as automatic comparison with oracle positions in low-level planning. We find that LLMs are strong at interpreting the high-level body movements but struggle with precise body part positioning. While breaking down motion queries into atomic components improves planning performance, LLMs have difficulty with multi-step movements involving high-degree-of-freedom body parts. Furthermore, LLMs provide reasonable approximation for general spatial descriptions, but fail to handle precise spatial specifications in text, and the precise spatial-temporal parameters needed for avatar control. Notably, LLMs show promise in conceptualizing creative motions and distinguishing culturally-specific motion patterns.

📄 PDF Abstract BibTeX arXiv:2505.21531

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Imitation versus Innovation: What children can do that large language and language-and-vision models cannot (yet)?

2023-05-08 · Eunice Yiu, Eliza Kosoy, Alison Gopnik

Much discussion about large language models and language-and-vision models has focused on whether these models are intelligent agents. We present an alternative perspective. We argue that these artificial intelligence mo…

Instructed to Bias: Instruction-Tuned Language Models Exhibit Emergent Cognitive Bias

2023-08-01 · Itay Itzhak, Gabriel Stanovsky, Nir Rosenfeld, Yonatan Belinkov

Recent studies show that instruction tuning (IT) and reinforcement learning from human feedback (RLHF) improve the abilities of large language models (LMs) dramatically. While these tuning methods can help align models w…

Decision Making

Language Statistics and False Belief Reasoning: Evidence from 41 Open-Weight LMs

2026-02-17 · Sean Trott, Samuel Taylor, Cameron Jones, James A. Michaelov 외 arxiv

Research on mental state reasoning in language models (LMs) has the potential to inform theories of human social cognition--such as the theory that mental state reasoning emerges in part from language exposure--and our u…

JECC: Commonsense Reasoning Tasks Derived from Interactive Fictions

2022-10-18 · Mo Yu, Yi Gu, Xiaoxiao Guo, Yufei Feng 외

Commonsense reasoning simulates the human ability to make presumptions about our physical world, and it is an essential cornerstone in building general AI systems. We propose a new commonsense reasoning dataset based on …

Reading Comprehension

Grounding Language about Belief in a Bayesian Theory-of-Mind

2024-02-16 · Lance Ying, Tan Zhi-Xuan, Lionel Wong, Vikash Mansinghka 외

Despite the fact that beliefs are mental states that cannot be directly observed, humans talk about each others' beliefs on a regular basis, often using rich compositional language to describe what others think and know.…

Attribute