paper-with-me

홈 › Papers

DanceEditor: Towards Iterative Editable Music-driven Dance Generation with Open-Vocabulary Descriptions

2025-08-24 · Hengyuan Zhang, Zhe Li, Xingqun Qi, Mengze Li, Muyi Sun, Man Zhang, Sirui Han arxiv

Generating coherent and diverse human dances from music signals has gained tremendous progress in animating virtual avatars. While existing methods support direct dance synthesis, they fail to recognize that enabling users to edit dance movements is far more practical in real-world choreography scenarios. Moreover, the lack of high-quality dance datasets incorporating iterative editing also limits addressing this challenge. To achieve this goal, we first construct DanceRemix, a large-scale multi-turn editable dance dataset comprising the prompt featuring over 25.3M dance frames and 84.5K pairs. In addition, we propose a novel framework for iterative and editable dance generation coherently aligned with given music signals, namely DanceEditor. Considering the dance motion should be both musical rhythmic and enable iterative editing by user descriptions, our framework is built upon a prediction-then-editing paradigm unifying multi-modal conditions. At the initial prediction stage, our framework improves the authority of generated results by directly modeling dance movements from tailored, aligned music. Moreover, at the subsequent iterative editing stages, we incorporate text descriptions as conditioning information to draw the editable results through a specifically designed Cross-modality Editing Module (CEM). Specifically, CEM adaptively integrates the initial prediction with music and text prompts as temporal motion cues to guide the synthesized sequences. Thereby, the results display music harmonics while preserving fine-grained semantic alignment with text descriptions. Extensive experiments demonstrate that our method outperforms the state-of-the-art models on our newly collected DanceRemix dataset. Code is available at https://lzvsdy.github.io/DanceEditor/.

📄 PDF Abstract BibTeX arXiv:2508.17342

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

EDGE: Editable Dance Generation From Music

2022-11-19 · CVPR 2023 1 · Jonathan Tseng, Rodrigo Castellon, C. Karen Liu

Dance is an important human art form, but creating new dances can be difficult and time-consuming. In this work, we introduce Editable Dance GEneration (EDGE), a state-of-the-art method for editable dance generation that…

DiversityMotion Synthesis

Text Dictates, Music Decorates: Energy-based Attention for Editable Dance Motion Generation

2026-06-22 · Seong Jong Yoo, Siyuan Peng, Felix Gu, Stratis Aloimonos 외 arxiv

Choreographic motion generation poses unique challenges for AI, demanding precise semantic control over complex, temporally structured, and expressive full-body dynamics. While existing models can synthesize motion from …

Libretto: Giving LLM Agents a Sense of Musical Structure

2026-06-21 · Yichen Xu arxiv

Generative music systems can now produce impressive audio from text prompts, but audio outputs are difficult to inspect, edit, and diagnose as musical structure. We introduce Libretto, an agent-facing framework for symbo…

Music Generation

DuetGen: Music Driven Two-Person Dance Generation via Hierarchical Masked Modeling

2025-06-23 · Anindita Ghosh, Bing Zhou, Rishabh Dabral, Jian Wang 외

We present DuetGen, a novel framework for generating interactive two-person dances from music. The key challenge of this task lies in the inherent complexities of two-person dance interactions, where the partners need to…

Motion Synthesis

Music-Driven Group Choreography

2023-03-22 · CVPR 2023 1 · Nhat Le, Thang Pham, Tuong Do, Erman Tjiputra 외

Music-driven choreography is a challenging problem with a wide variety of industrial applications. Recently, many methods have been proposed to synthesize dance motions from music for a single dancer. However, generating…

Motion Synthesis