paper-with-me

Music Generation

1개 벤치마크 · 논문 481편 · 이 태스크의 논문 보기 →

Benchmarks

Most implemented

Music Transformer

2018-09-12 · 구현 12개

Simple and Controllable Music Generation

2023-06-08 · 구현 8개

MusicLM: Generating Music From Text

2023-01-26 · 구현 5개

Papers

VIBE: Video Instruction-aligned Background music gEneration

2026-08-31 · Aryan Vijay Bhosale, Vaibhavi Lokegaonkar, Vishnu Raj, Gouthaman KV 외 arxiv

Current video-to-music (V2M) models lack semantic control and fail to penalize instruction violations, largely due to their reliance on reconstruction objectives and the representational bottleneck of static cross-modal …

Instruction FollowingMusic Generation

Learning Music Style for Piano Arrangement Through Cross-Modal Bootstrapping

2026-08-04 · Jingwei Zhao, Gus Xia, Ziyu Wang, Ye Wang arxiv

What is music style? Though often described using text labels such as "swing," "classical," or "emotional," the real style remains implicit and hidden in concrete music examples. In this paper, we introduce a cross-modal…

Contrastive LearningMusic GenerationStyle Transfer

Pushing the Frontier of Full-Song Generation: Hierarchical Autoregressive Planning Meets Flow-Matching Rendering

2026-07-22 · Junyu Dai, Xinyue Fan, Weiqin Li, Xiangang Li 외 arxiv

In this report, we present a unified song generation framework capable of producing high-quality full-length music from lyrics, text descriptions, and musical attributes. The proposed framework supports three tasks: Lyri…

Music Generation

RPPNet: Perceptually-Grouped Rhythm-Pitch Primitives for Long-Term Structure Melody Generation via Boundary-Aware Modeling

2026-07-22 · Tieyao Zhang, Yuke Liu, Jiaxing Yu, Xinda Wu 외 arxiv

Existing symbolic music generation models typically use bars as the basic structural unit. However, human perception of musical phrases often does not align with notated bar lines, leading to long-term structural fragmen…

Music Generation

WanSong v1.0 Technical Report

2026-07-16 · Binghui Chen, Pandeng Li, Yu Liu, Jingren Zhou arxiv

Music generation foundation models have recently attracted significant industry attention. However, achieving efficient generation and high-fidelity long-form audio while supporting controllability remains challenging. T…

Music Generation

Qwen-Music Technical Report

2026-07-13 · Jin Xu, Kangdi Wang, Ruibin Yuan, Shun Lei 외 hf

In this report, we introduce Qwen-Music, a powerful music generation model capable of producing highly musical and high-fidelity songs with complete vocal singing. Qwen-Music supports two core tasks: Text to Music Genera…

Music Generation

전체 481편 보기 →