paper-with-me

홈 › Papers

Equipping Pretrained Unconditional Music Transformers with Instrument and Genre Controls

2023-11-21 · Weihan Xu, Julian McAuley, Shlomo Dubnov, Hao-Wen Dong

The ''pretraining-and-finetuning'' paradigm has become a norm for training domain-specific models in natural language processing and computer vision. In this work, we aim to examine this paradigm for symbolic music generation through leveraging the largest ever symbolic music dataset sourced from the MuseScore forum. We first pretrain a large unconditional transformer model using 1.5 million songs. We then propose a simple technique to equip this pretrained unconditional music transformer model with instrument and genre controls by finetuning the model with additional control tokens. Our proposed representation offers improved high-level controllability and expressiveness against two existing representations. The experimental results show that the proposed model can successfully generate music with user-specified instruments and genre. In a subjective listening test, the proposed model outperforms the pretrained baseline model in terms of coherence, harmony, arrangement and overall quality.

📄 PDF Abstract BibTeX arXiv:2311.12257

Code (0)

등록된 구현이 없습니다.

Tasks

Music Generation

Similar Papers 제목 키워드 기반

Diff-TONE: Timestep Optimization for iNstrument Editing in Text-to-Music Diffusion Models

2025-06-18 · Teysir Baoueb, Xiaoyu Bie, Xi Wang, Gaël Richard

Breakthroughs in text-to-music generation models are transforming the creative landscape, equipping musicians with innovative tools for composition and experimentation like never before. However, controlling the generati…

Music GenerationText-to-Music Generation

SingSong: Generating musical accompaniments from singing

2023-01-30 · Chris Donahue, Antoine Caillon, Adam Roberts, Ethan Manilow 외

We present SingSong, a system that generates instrumental music to accompany input vocals, potentially offering musicians and non-musicians alike an intuitive new way to create music featuring their own voice. To accompl…

Audio GenerationRetrieval

Music Tempo Estimation on Solo Instrumental Performance

2025-04-25 · Zhanhong He, Roberto Togneri, Xiangyu Zhang

Recently, automatic music transcription has made it possible to convert musical audio into accurate MIDI. However, the resulting MIDI lacks music notations such as tempo, which hinders its conversion into sheet music. In…

Music Transcription

Transformer-Based Approaches for Automatic Music Transcription

2021-02-12 · Conference 2021 2 · Christos Zonios

Automatic Music Transcription (AMT) is the process of extracting information from audio into some form of music notation. In polyphonic music, this is a very hard problem for computers to solve as it requires significa…

Language ModellingMusic Transcriptionspeech-recognitionSpeech Recognition+1

GTR-CTRL: Instrument and Genre Conditioning for Guitar-Focused Music Generation with Transformers

2023-02-10 · Pedro Sarmento, Adarsh Kumar, Yu-Hua Chen, CJ Carr 외

Recently, symbolic music generation with deep learning techniques has witnessed steady improvements. Most works on this topic focus on MIDI representations, but less attention has been paid to symbolic music generation u…

Genre classificationMusic GenerationRhythm