paper-with-me

Papers

Enhanced Fine-grained Motion Diffusion for Text-driven Human Motion Synthesis

2023-05-23 · Dong Wei, Xiaoning Sun, Huaijiang Sun, Bin Li, Shengxiang Hu, Weiqing Li, Jianfeng Lu

The emergence of text-driven motion synthesis technique provides animators with great potential to create efficiently. However, in most cases, textual expressions only contain general and qualitative motion descriptions, while lack fine depiction and sufficient intensity, leading to the synthesized motions that either (a) semantically compliant but uncontrollable over specific pose details, or (b) even deviates from the provided descriptions, bringing animators with undesired cases. In this paper, we propose DiffKFC, a conditional diffusion model for text-driven motion synthesis with KeyFrames Collaborated, enabling realistic generation with collaborative and efficient dual-level control: coarse guidance at semantic level, with only few keyframes for direct and fine-grained depiction down to body posture level. Unlike existing inference-editing diffusion models that incorporate conditions without training, our conditional diffusion model is explicitly trained and can fully exploit correlations among texts, keyframes and the diffused target frames. To preserve the control capability of discrete and sparse keyframes, we customize dilated mask attention modules where only partial valid tokens participate in local-to-global attention, indicated by the dilated keyframe mask. Additionally, we develop a simple yet effective smoothness prior, which steers the generated frames towards seamless keyframe transitions at inference. Extensive experiments show that our model not only achieves state-of-the-art performance in terms of semantic fidelity, but more importantly, is able to satisfy animator requirements through fine-grained guidance without tedious labor.

📄 PDF Abstract BibTeX arXiv:2305.13773

Code (0)

등록된 구현이 없습니다.

Tasks

Motion Synthesisvalid

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Free-T2M: Frequency Enhanced Text-to-Motion Diffusion Model With Consistency Loss

2025-01-30 · Wenshuo Chen, Haozhe Jia, Songning Lai, Keming Wu 외

Rapid progress in text-to-motion generation has been largely driven by diffusion models. However, existing methods focus solely on temporal modeling, thereby overlooking frequency-domain analysis. We identify two key pha…

DenoisingMotion GenerationMotion Synthesis

AnimateAnything: Fine-Grained Open Domain Image Animation with Motion Guidance

2023-11-21 · Zuozhuo Dai, Zhenghao Zhang, Yao Yao, Bingxue Qiu 외

Image animation is a key task in computer vision which aims to generate dynamic visual content from static image. Recent image animation methods employ neural based rendering technique to generate realistic animations. D…

Image AnimationImage to Video GenerationVideo Generation

FG-MDM: Towards Zero-Shot Human Motion Generation via ChatGPT-Refined Descriptions

2023-12-05 · Xu Shi, Wei Yao, Chuanchen Luo, Junran Peng 외

Recently, significant progress has been made in text-based motion generation, enabling the generation of diverse and high-quality human motions that conform to textual descriptions. However, generating motions beyond the…

Language ModelingLanguage ModellingLarge Language ModelMotion Generation

Animus3D: Text-driven 3D Animation via Motion Score Distillation

2025-12-14 · Qi Sun, Can Wang, Jiaxiang Shang, Wensen Feng 외 arxiv

We present Animus3D, a text-driven 3D animation framework that generates motion field given a static 3D asset and text prompt. Previous methods mostly leverage the vanilla Score Distillation Sampling (SDS) objective to d…

Noise Estimation

FineMoGen: Fine-Grained Spatio-Temporal Motion Generation and Editing

2023-12-22 · NeurIPS 2023 11 · Mingyuan Zhang, Huirong Li, Zhongang Cai, Jiawei Ren 외

Text-driven motion generation has achieved substantial progress with the emergence of diffusion models. However, existing methods still struggle to generate complex motion sequences that correspond to fine-grained descri…

Mixture-of-ExpertsMotion GenerationMotion Synthesis