paper-with-me

Papers

SingSong: Generating musical accompaniments from singing

2023-01-30 · Chris Donahue, Antoine Caillon, Adam Roberts, Ethan Manilow, Philippe Esling, Andrea Agostinelli, Mauro Verzetti, Ian Simon, Olivier Pietquin, Neil Zeghidour, Jesse Engel

We present SingSong, a system that generates instrumental music to accompany input vocals, potentially offering musicians and non-musicians alike an intuitive new way to create music featuring their own voice. To accomplish this, we build on recent developments in musical source separation and audio generation. Specifically, we apply a state-of-the-art source separation algorithm to a large corpus of music audio to produce aligned pairs of vocals and instrumental sources. Then, we adapt AudioLM (Borsos et al., 2022) -- a state-of-the-art approach for unconditional audio generation -- to be suitable for conditional "audio-to-audio" generation tasks, and train it on the source-separated (vocal, instrumental) pairs. In a pairwise comparison with the same vocal inputs, listeners expressed a significant preference for instrumentals generated by SingSong compared to those from a strong retrieval baseline. Sound examples at https://g.co/magenta/singsong

📄 PDF Abstract BibTeX arXiv:2301.12662

Code (0)

등록된 구현이 없습니다.

Tasks

Audio GenerationRetrieval

Similar Papers 제목 키워드 기반

FastSAG: Towards Fast Non-Autoregressive Singing Accompaniment Generation

2024-05-13 · Jianyi Chen, Wei Xue, Xu Tan, Zhen Ye 외

Singing Accompaniment Generation (SAG), which generates instrumental music to accompany input vocals, is crucial to developing human-AI symbiotic art creation systems. The state-of-the-art method, SingSong, utilizes a mu…

Rhythm

Sing-On-Your-Beat: Simple Text-Controllable Accompaniment Generations

2024-11-03 · Quoc-Huy Trinh, Minh-Van Nguyen, Trong-Hieu Nguyen Mau, Khoa Tran 외

Singing is one of the most cherished forms of human entertainment. However, creating a beautiful song requires an accompaniment that complements the vocals and aligns well with the song instruments and genre. With advanc…

Generalized Multi-Source Inference for Text Conditioned Music Diffusion Models

2024-03-18 · Emilian Postolache, Giorgio Mariani, Luca Cosmo, Emmanouil Benetos 외

Multi-Source Diffusion Models (MSDM) allow for compositional musical generation tasks: generating a set of coherent sources, creating accompaniments, and performing source separation. Despite their versatility, they requ…

Synchronising speech segments with musical beats in Mandarin and English singing

2021-06-18 · Cong Zhang, Jian Zhu

Generating synthesised singing voice with models trained on speech data has many advantages due to the models' flexibility and controllability. However, since the information about the temporal relationship between segme…

Versatile Framework for Song Generation with Prompt-based Control

2025-04-27 · Yu Zhang, Wenxiang Guo, Changhao Pan, Zhiyuan Zhu 외

Song generation focuses on producing controllable high-quality songs based on various prompts. However, existing methods struggle to generate vocals and accompaniments with prompt-based control and proper alignment. Addi…