paper-with-me

홈 › Papers

MotionGPT: Finetuned LLMs Are General-Purpose Motion Generators

2023-06-19 · Yaqi Zhang, Di Huang, Bin Liu, Shixiang Tang, Yan Lu, Lu Chen, Lei Bai, Qi Chu, Nenghai Yu, Wanli Ouyang

Generating realistic human motion from given action descriptions has experienced significant advancements because of the emerging requirement of digital humans. While recent works have achieved impressive results in generating motion directly from textual action descriptions, they often support only a single modality of the control signal, which limits their application in the real digital human industry. This paper presents a Motion General-Purpose generaTor (MotionGPT) that can use multimodal control signals, e.g., text and single-frame poses, for generating consecutive human motions by treating multimodal signals as special input tokens in large language models (LLMs). Specifically, we first quantize multimodal control signals into discrete codes and then formulate them in a unified prompt instruction to ask the LLMs to generate the motion answer. Our MotionGPT demonstrates a unified human motion generation model with multimodal control signals by tuning a mere 0.4% of LLM parameters. To the best of our knowledge, MotionGPT is the first method to generate human motion by multimodal control signals, which we hope can shed light on this new direction. Visit our webpage at https://qiqiapink.github.io/MotionGPT/.

📄 PDF Abstract BibTeX arXiv:2306.10900

Code (0)

등록된 구현이 없습니다.

Tasks

Motion Generation

Similar Papers 제목 키워드 기반

MotionGPT-2: A General-Purpose Motion-Language Model for Motion Generation and Understanding

2024-10-29 · YuAn Wang, Di Huang, Yaqi Zhang, Wanli Ouyang 외

Generating lifelike human motions from descriptive texts has experienced remarkable research focus in the recent years, propelled by the emerging requirements of digital humans.Despite impressive advances, existing appro…

DescriptiveLanguage ModelingLanguage ModellingMotion Captioning+2

MotionGPT: Human Motion as a Foreign Language

2023-06-26 · NeurIPS 2023 11 · Biao Jiang, Xin Chen, Wen Liu, Jingyi Yu 외

Though the advancement of pre-trained large language models unfolds, the exploration of building a unified model for language and other multi-modal data, such as motion, remains challenging and untouched so far. Fortunat…

Language ModelingLanguage ModellingMotion CaptioningMotion Generation+3

From Diffusion to Flow: Efficient Motion Generation in MotionGPT3

2026-03-23 · Jaymin Ban, JiHong Jeon, SangYeop Jeong arxiv

Recent text-driven motion generation methods span both discrete token-based approaches and continuous-latent formulations. MotionGPT3 exemplifies the latter paradigm, combining a learned continuous motion latent space wi…

Audio Generation

ECG-LLM -- training and evaluation of domain-specific large language models for electrocardiography

2025-10-21 · Lara Ahrens, Wilhelm Haverkamp, Nils Strodthoff arxiv

Domain-adapted open-weight large language models (LLMs) offer promising healthcare applications, from queryable knowledge bases to multimodal assistants, with the crucial advantage of local deployment for privacy preserv…

PRISMA-DFLLM: An Extension of PRISMA for Systematic Literature Reviews using Domain-specific Finetuned Large Language Models

2023-06-15 · Teo Susnjak

With the proliferation of open-sourced Large Language Models (LLMs) and efficient finetuning techniques, we are on the cusp of the emergence of numerous domain-specific LLMs that have been finetuned for expertise across …