paper-with-me

홈 › Papers

Unimotion: Unifying 3D Human Motion Synthesis and Understanding

2024-09-24 · Chuqiao Li, Julian Chibane, Yannan He, Naama Pearl, Andreas Geiger, Gerard Pons-Moll

We introduce Unimotion, the first unified multi-task human motion model capable of both flexible motion control and frame-level motion understanding. While existing works control avatar motion with global text conditioning, or with fine-grained per frame scripts, none can do both at once. In addition, none of the existing works can output frame-level text paired with the generated poses. In contrast, Unimotion allows to control motion with global text, or local frame-level text, or both at once, providing more flexible control for users. Importantly, Unimotion is the first model which by design outputs local text paired with the generated poses, allowing users to know what motion happens and when, which is necessary for a wide range of applications. We show Unimotion opens up new applications: 1.) Hierarchical control, allowing users to specify motion at different levels of detail, 2.) Obtaining motion text descriptions for existing MoCap data or YouTube videos 3.) Allowing for editability, generating motion from text, and editing the motion via text edits. Moreover, Unimotion attains state-of-the-art results for the frame-level text-to-motion task on the established HumanML3D dataset. The pre-trained model and code are available available on our project page at https://coral79.github.io/uni-motion/.

📄 PDF Abstract BibTeX arXiv:2409.15904

Code (0)

등록된 구현이 없습니다.

Tasks

Motion Synthesis

Similar Papers 제목 키워드 기반

UniMotion: A Unified Framework for Motion-Text-Vision Understanding and Generation

2026-03-23 · Ziyi Wang, Xinshun Wang, Shuang Chen, Yang Cong 외 arxiv

We present UniMotion, to our knowledge the first unified framework for simultaneous understanding and generation of human motion, natural language, and RGB images within a single architecture. Existing unified models han…

UniMotion: A Unified Motion Framework for Simulation, Prediction and Planning

2026-01-31 · Nan Song, Junzhe Jiang, Jingyu Li, Xiatian Zhu 외 arxiv

Motion simulation, prediction and planning are foundational tasks in autonomous driving, each essential for modeling and reasoning about dynamic traffic scenarios. While often addressed in isolation due to their differin…

Autonomous Driving

X-UniMotion: Animating Human Images with Expressive, Unified and Identity-Agnostic Motion Latents

2025-08-12 · Guoxian Song, Hongyi Xu, Xiaochen Zhao, You Xie 외 arxiv

We present X-UniMotion, a unified and expressive implicit latent representation for whole-body human motion, encompassing facial expressions, body poses, and hand gestures. Unlike prior motion transfer methods that rely …

AvatarGPT: All-in-One Framework for Motion Understanding Planning Generation and Beyond

2024-01-01 · CVPR 2024 1 · Zixiang Zhou, Yu Wan, Baoyuan Wang

Large Language Models(LLMs) have shown remarkable emergent abilities in unifying almost all (if not every) NLP tasks. In the human motion-related realm however researchers still develop siloed models for each task. I…

AllMotion Synthesis

AvatarGPT: All-in-One Framework for Motion Understanding, Planning, Generation and Beyond

2023-11-28 · Zixiang Zhou, Yu Wan, Baoyuan Wang

Large Language Models(LLMs) have shown remarkable emergent abilities in unifying almost all (if not every) NLP tasks. In the human motion-related realm, however, researchers still develop siloed models for each task. Ins…

AllMotion Synthesis