paper-with-me

홈 › Papers

Benchmarking Music Generation Models and Metrics via Human Preference Studies

2025-06-23 · Audio Imagination: NeurIPS 2024 Workshop 2024 10 · Florian Grötschla, Ahmet Solak, Luca A. Lanzendörfer, Roger Wattenhofer

Recent advancements have brought generated music closer to human-created compositions, yet evaluating these models remains challenging. While human preference is the gold standard for assessing quality, translating these subjective judgments into objective metrics, particularly for text-audio alignment and music quality, has proven difficult. In this work, we generate 6k songs using 12 state-of-the-art models and conduct a survey of 15k pairwise audio comparisons with 2.5k human participants to evaluate the correlation between human preferences and widely used metrics. To the best of our knowledge, this work is the first to rank current state-of-the-art music generation models and metrics based on human preference. To further the field of subjective metric evaluation, we provide open access to our dataset of generated music and human evaluations.

📄 PDF Abstract BibTeX arXiv:2506.19085

Code (0)

등록된 구현이 없습니다.

Tasks

BenchmarkingMusic Generation

Similar Papers 제목 키워드 기반

From Aesthetics to Human Preferences: Comparative Perspectives of Evaluating Text-to-Music Systems

2025-04-30 · huan zhang, Jinhua Liang, Huy Phan, Wenwu Wang 외

Evaluating generative models remains a fundamental challenge, particularly when the goal is to reflect human preferences. In this paper, we use music generation as a case study to investigate the gap between automatic ev…

Music Generation

Aligning Generative Music AI with Human Preferences: Methods and Challenges

2025-11-19 · Dorien Herremans, Abhinaba Roy arxiv

Recent advances in generative AI for music have achieved remarkable fidelity and stylistic diversity, yet these systems often fail to align with nuanced human preferences due to the specific loss functions they use. This…

Music Generation

Aligning Text-to-Music Evaluation with Human Preferences

2025-03-20 · Yichen Huang, Zachary Novack, Koichi Saito, Jiatong Shi 외

Despite significant recent advances in generative acoustic text-to-music (TTM) modeling, robust evaluation of these models lags behind, relying in particular on the popular Fr\'echet Audio Distance (FAD). In this work, w…

FAD

MusicRL: Aligning Music Generation to Human Preferences

2024-02-06 · Geoffrey Cideron, Sertan Girgin, Mauro Verzetti, Damien Vincent 외

We propose MusicRL, the first music generation system finetuned from human feedback. Appreciation of text-to-music models is particularly subjective since the concept of musicality as well as the specific intention behin…

Music Generation

LeVo: High-Quality Song Generation with Multi-Preference Alignment

2025-06-09 · Shun Lei, Yaoxun Xu, Zhiwei Lin, Huaicheng Zhang 외

Recent advances in large language models (LLMs) and audio language models have significantly improved music generation, particularly in lyrics-to-song generation. However, existing approaches still struggle with the comp…

Instruction FollowingMusic Generation