paper-with-me

홈 › Papers

Voice and accompaniment separation in music using self-attention convolutional neural network

2020-03-19 · Yuzhou Liu, Balaji Thoshkahna, Ali Milani, Trausti Kristjansson

Music source separation has been a popular topic in signal processing for decades, not only because of its technical difficulty, but also due to its importance to many commercial applications, such as automatic karoake and remixing. In this work, we propose a novel self-attention network to separate voice and accompaniment in music. First, a convolutional neural network (CNN) with densely-connected CNN blocks is built as our base network. We then insert self-attention subnets at different levels of the base CNN to make use of the long-term intra-dependency of music, i.e., repetition. Within self-attention subnets, repetitions of the same musical patterns inform reconstruction of other repetitions, for better source separation performance. Results show the proposed method leads to 19.5% relative improvement in vocals separation in terms of SDR. We compare our methods with state-of-the-art systems i.e. MMDenseNet and MMDenseLSTM.

📄 PDF Abstract BibTeX arXiv:2003.08954

Code (0)

등록된 구현이 없습니다.

Tasks

Music Source Separation

Similar Papers 제목 키워드 기반

Investigation of Singing Voice Separation for Singing Voice Detection in Polyphonic Music

2020-04-08 · Yifu Sun, xulong Zhang, Yi Yu, Xi Chen 외

Singing voice detection (SVD), to recognize vocal parts in the song, is an essential task in music information retrieval (MIR). The task remains challenging since singing voice varies and intertwines with the accompanime…

Information RetrievalMelody ExtractionMusic Information RetrievalRetrieval

Singing Beat Tracking With Self-supervised Front-end and Linear Transformers

2022-08-31 · Mojtaba Heydari, Zhiyao Duan

Tracking beats of singing voices without the presence of musical accompaniment can find many applications in music production, automatic song arrangement, and social media interaction. Its main challenge is the lack of s…

Beat Tracking

Informed Group-Sparse Representation for Singing Voice Separation

2018-01-09 · Tak-Shing T. Chan, Yi-Hsuan Yang

Singing voice separation attempts to separate the vocal and instrumental parts of a music recording, which is a fundamental problem in music information retrieval. Recent work on singing voice separation has shown that t…

Information RetrievalMusic Information RetrievalRetrieval

Improved Speech Enhancement with the Wave-U-Net

2018-11-27 · Craig Macartney, Tillman Weyde

We study the use of the Wave-U-Net architecture for speech enhancement, a model introduced by Stoller et al for the separation of music vocals and accompaniment. This end-to-end learning method for audio source separatio…

Audio Source SeparationSpeech Enhancementspeech-recognitionSpeech Recognition

Improved Speech Enhancement with the Wave-U-Net

2018-10-22 · Anonymous

We study the use of the Wave-U-Net architecture for speech enhancement, a model introduced by Stoller et al for the separation of music vocals and accompaniment. This end-to-end learning method for audio source separati…

Audio Source SeparationSpeech Enhancementspeech-recognitionSpeech Recognition