paper-with-me

Papers

Bidirectional Autoregessive Diffusion Model for Dance Generation

2024-01-01 · CVPR 2024 1 · Canyu Zhang, YouBao Tang, Ning Zhang, Ruei-Sung Lin, Mei Han, Jing Xiao, Song Wang

Dance serves as a powerful medium for expressing human emotions but the lifelike generation of dance is still a considerable challenge. Recently diffusion models have showcased remarkable generative abilities across various domains. They hold promise for human motion generation due to their adaptable many-to-many nature. Nonetheless current diffusion-based motion generation models often create entire motion sequences directly and unidirectionally lacking focus on the motion with local and bidirectional enhancement. When choreographing high-quality dance movements people need to take into account not only the musical context but also the nearby music-aligned dance motions. To authentically capture human behavior we propose a Bidirectional Autoregressive Diffusion Model (BADM) for music-to-dance generation where a bidirectional encoder is built to enforce that the generated dance is harmonious in both the forward and backward directions. To make the generated dance motion smoother a local information decoder is built for local motion enhancement. The proposed framework is able to generate new motions based on the input conditions and nearby motions which foresees individual motion slices iteratively and consolidates all predictions. To further refine the synchronicity between the generated dance and the beat the beat information is incorporated as an input to generate better music-aligned dance movements. Experimental results demonstrate that the proposed model achieves state-of-the-art performance compared to existing unidirectional approaches on the prominent benchmark for music-to-dance generation.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

modelMotion Generation

Methods 이 논문이 사용한 방법론

Focus 설명 없음
Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Bidirectional Autoregressive Diffusion Model for Dance Generation

2024-02-06 · Canyu Zhang, YouBao Tang, Ning Zhang, Ruei-Sung Lin 외

Dance serves as a powerful medium for expressing human emotions, but the lifelike generation of dance is still a considerable challenge. Recently, diffusion models have showcased remarkable generative abilities across va…

modelMotion Generation

BatGPT: A Bidirectional Autoregessive Talker from Generative Pre-trained Transformer

2023-07-01 · Zuchao Li, Shitou Zhang, Hai Zhao, Yifei Yang 외

BatGPT is a large-scale language model designed and trained jointly by Wuhan University and Shanghai Jiao Tong University. It is capable of generating highly natural and fluent text in response to various types of input,…

Language ModelingLanguage ModellingQuestion AnsweringText Generation

Text-to-3D Generation with Bidirectional Diffusion using both 2D and 3D priors

2023-12-07 · CVPR 2024 1 · Lihe Ding, Shaocong Dong, Zhanpeng Huang, Zibin Wang 외

Most 3D generation research focuses on up-projecting 2D foundation models into the 3D space, either by minimizing 2D Score Distillation Sampling (SDS) loss or fine-tuning on multi-view datasets. Without explicit 3D prior…

3D GenerationDiversityText to 3DTexture Synthesis

ST-GDance++: A Scalable Spatial-Temporal Diffusion for Long-Duration Group Choreography

2026-03-20 · Jing Xu, Weiqiang Wang, Cunjian Chen, Jun Liu 외 arxiv

Group dance generation from music requires synchronizing multiple dancers while maintaining spatial coordination, making it highly relevant to applications such as film production, gaming, and animation. Recent group dan…

Eulerian Motion Guidance: Robust Image Animation via Bidirectional Geometric Consistency

2026-05-07 · Thong Nguyen, Khoi M. Le, Cong-Duy Nguyen, Luu Anh Tuan 외 arxiv

Recent advancements in image animation have utilized diffusion models to breathe life into static images. However, existing controllable frameworks typically rely on Lagrangian motion guidance, where optical flow is esti…