paper-with-me

Papers

MDD: A Dataset for Text-and-Music Conditioned Duet Dance Generation

2025-08-23 · Prerit Gupta, Jason Alexander Fotso-Puepi, Zhengyuan Li, Jay Mehta, Aniket Bera arxiv

We introduce Multimodal DuetDance (MDD), a diverse multimodal benchmark dataset designed for text-controlled and music-conditioned 3D duet dance motion generation. Our dataset comprises 620 minutes of high-quality motion capture data performed by professional dancers, synchronized with music, and detailed with over 10K fine-grained natural language descriptions. The annotations capture a rich movement vocabulary, detailing spatial relationships, body movements, and rhythm, making MDD the first dataset to seamlessly integrate human motions, music, and text for duet dance generation. We introduce two novel tasks supported by our dataset: (1) Text-to-Duet, where given music and a textual prompt, both the leader and follower dance motion are generated (2) Text-to-Dance Accompaniment, where given music, textual prompt, and the leader's motion, the follower's motion is generated in a cohesive, text-aligned manner. We include baseline evaluations on both tasks to support future research.

📄 PDF Abstract BibTeX arXiv:2508.16911

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

DuetGen: Music Driven Two-Person Dance Generation via Hierarchical Masked Modeling

2025-06-23 · Anindita Ghosh, Bing Zhou, Rishabh Dabral, Jian Wang 외

We present DuetGen, a novel framework for generating interactive two-person dances from music. The key challenge of this task lies in the inherent complexities of two-person dance interactions, where the partners need to…

Motion Synthesis

Duolando: Follower GPT with Off-Policy Reinforcement Learning for Dance Accompaniment

2024-03-27 · Li SiYao, Tianpei Gu, Zhitao Yang, Zhengyu Lin 외

We introduce a novel task within the field of 3D dance generation, termed dance accompaniment, which necessitates the generation of responsive movements from a dance partner, the "follower", synchronized with the lead da…

Rhythm

TeMuDance: Contrastive Alignment-Based Textual Control for Music-Driven Dance Generation

2026-04-18 · Xinran Liu, Diptesh Kanojia, Wenwu Wang, Zhenhua Feng arxiv

Existing music-driven dance generation approaches have achieved strong realism and effective audio-motion alignment. However, they generally lack semantic controllability, making it difficult to guide specific movements …

Cross-Modal Retrieval

ReactDance: Progressive-Granular Representation for Long-Term Coherent Reactive Dance Generation

2025-05-08 · Jingzhong Lin, Yuanyuan Qi, Xinru Li, Wenxuan Huang 외

Reactive dance generation (RDG) produces follower movements conditioned on guiding dancer and music while ensuring spatial coordination and temporal coherence. However, existing methods overemphasize global constraints a…

Quantization

Transflower: probabilistic autoregressive dance generation with multimodal attention

2021-06-25 · Guillermo Valle-Pérez, Gustav Eje Henter, Jonas Beskow, André Holzapfel 외

Dance requires skillful composition of complex movements that follow rhythmic, tonal and timbral features of music. Formally, generating dance conditioned on a piece of music can be expressed as a problem of modelling a …