paper-with-me

홈 › Papers

HUMANISE: Language-conditioned Human Motion Generation in 3D Scenes

2022-10-18 · Zan Wang, Yixin Chen, Tengyu Liu, Yixin Zhu, Wei Liang, Siyuan Huang

Learning to generate diverse scene-aware and goal-oriented human motions in 3D scenes remains challenging due to the mediocre characteristics of the existing datasets on Human-Scene Interaction (HSI); they only have limited scale/quality and lack semantics. To fill in the gap, we propose a large-scale and semantic-rich synthetic HSI dataset, denoted as HUMANISE, by aligning the captured human motion sequences with various 3D indoor scenes. We automatically annotate the aligned motions with language descriptions that depict the action and the unique interacting objects in the scene; e.g., sit on the armchair near the desk. HUMANISE thus enables a new generation task, language-conditioned human motion generation in 3D scenes. The proposed task is challenging as it requires joint modeling of the 3D scene, human motion, and natural language. To tackle this task, we present a novel scene-and-language conditioned generative model that can produce 3D human motions of the desirable action interacting with the specified objects. Our experiments demonstrate that our model generates diverse and semantically consistent human motions in 3D scenes.

📄 PDF Abstract BibTeX arXiv:2210.09729

Code (1)

Silverster98/HUMANISE 공식 구현 pytorch

Tasks

Motion Generation

Similar Papers 제목 키워드 기반

Move as You Say Interact as You Can: Language-guided Human Motion Generation with Scene Affordance

2024-01-01 · CVPR 2024 1 · Zan Wang, Yixin Chen, Baoxiong Jia, Puhao Li 외

Despite significant advancements in text-to-motion synthesis generating language-guided human motion within 3D environments poses substantial challenges. These challenges stem primarily from (i) the absence of powerf…

Motion GenerationMotion Synthesis

Move as You Say, Interact as You Can: Language-guided Human Motion Generation with Scene Affordance

2024-03-26 · Zan Wang, Yixin Chen, Baoxiong Jia, Puhao Li 외

Despite significant advancements in text-to-motion synthesis, generating language-guided human motion within 3D environments poses substantial challenges. These challenges stem primarily from (i) the absence of powerful …

Motion GenerationMotion Synthesis

GHOST: Grounded Human Motion Generation with Open Vocabulary Scene-and-Text Contexts

2024-04-08 · Zoltán Á. Milacski, Koichiro Niinuma, Ryosuke Kawamura, Fernando de la Torre 외

The connection between our 3D surroundings and the descriptive language that characterizes them would be well-suited for localizing and generating human motion in context but for one problem. The complexity introduced by…

DescriptiveImage SegmentationKnowledge DistillationMotion Generation+1

Data over dialogue: Why artificial intelligence is unlikely to humanise medicine

2025-04-10 · Joshua Hatherley

Recently, a growing number of experts in artificial intelligence (AI) and medicine have be-gun to suggest that the use of AI systems, particularly machine learning (ML) systems, is likely to humanise the practice of medi…

UniHM: Universal Human Motion Generation with Object Interactions in Indoor Scenes

2025-05-19 · Zichen Geng, Zeeshan Hayder, Wei Liu, Ajmal Mian

Human motion synthesis in complex scenes presents a fundamental challenge, extending beyond conventional Text-to-Motion tasks by requiring the integration of diverse modalities such as static environments, movable object…

Human-Object Interaction DetectionMotion GenerationMotion SynthesisQuantization