paper-with-me

홈 › Papers

MuseBarControl: Enhancing Fine-Grained Control in Symbolic Music Generation through Pre-Training and Counterfactual Loss

2024-07-05 · Yangyang Shu, HaiMing Xu, Ziqin Zhou, Anton Van Den Hengel, Lingqiao Liu

Automatically generating symbolic music-music scores tailored to specific human needs-can be highly beneficial for musicians and enthusiasts. Recent studies have shown promising results using extensive datasets and advanced transformer architectures. However, these state-of-the-art models generally offer only basic control over aspects like tempo and style for the entire composition, lacking the ability to manage finer details, such as control at the level of individual bars. While fine-tuning a pre-trained symbolic music generation model might seem like a straightforward method for achieving this finer control, our research indicates challenges in this approach. The model often fails to respond adequately to new, fine-grained bar-level control signals. To address this, we propose two innovative solutions. First, we introduce a pre-training task designed to link control signals directly with corresponding musical tokens, which helps in achieving a more effective initialization for subsequent fine-tuning. Second, we implement a novel counterfactual loss that promotes better alignment between the generated music and the control prompts. Together, these techniques significantly enhance our ability to control music generation at the bar level, showing a 13.06\% improvement over conventional methods. Our subjective evaluations also confirm that this enhanced control does not compromise the musical quality of the original pre-trained generative model.

📄 PDF Abstract BibTeX arXiv:2407.04331

Code (0)

등록된 구현이 없습니다.

Tasks

counterfactualMusic Generation

Similar Papers 제목 키워드 기반

FIGARO: Generating Symbolic Music with Fine-Grained Artistic Control

2022-01-26 · Dimitri von Rütte, Luca Biggio, Yannic Kilcher, Thomas Hofmann

Generating music with deep neural networks has been an area of active research in recent years. While the quality of generated samples has been steadily increasing, most methods are only able to exert minimal control ove…

Inductive BiasMusic Generation

E0: Enhancing Generalization and Fine-Grained Control in VLA Models via Tweedie Discrete Diffusion

2025-11-26 · Zhihao Zhan, Jiaying Zhou, Likui Zhang, Qinhan Lv 외 arxiv

Vision-Language-Action (VLA) models offer a unified framework for robotic manipulation by integrating visual perception, language understanding, and control generation. However, existing VLA systems still struggle to gen…

Efficient Fine-Grained Guidance for Diffusion Model Based Symbolic Music Generation

2024-10-11 · Tingyu Zhu, Haoyu Liu, Ziyu Wang, Zhimin Jiang 외

Developing generative models to create or conditionally create symbolic music presents unique challenges due to the combination of limited data availability and the need for high precision in note pitch. To address these…

Music Generation

Dissecting Logical Reasoning in LLMs: A Fine-Grained Evaluation and Supervision Study

2025-06-05 · Yujun Zhou, Jiayi Ye, Zipeng Ling, Yufei Han 외

Logical reasoning is a core capability for many applications of large language models (LLMs), yet existing benchmarks often rely solely on final-answer accuracy, failing to capture the quality and structure of the reason…

Logical Reasoning

Composer Vector: Style-steering Symbolic Music Generation in a Latent Space

2026-04-03 · Xunyi Jiang, Mingyang Yao, Jingyue Huang, Julian McAuley arxiv

Symbolic music generation has made significant progress, yet achieving fine-grained and flexible control over composer style remains challenging. Existing training-based methods for composer style conditioning depend on …

Music Generation