paper-with-me

Papers

Spectrogram-channels u-net: a source separation model viewing each channel as the spectrogram of each source

2018-10-26 · Jaehoon Oh, Duyeon Kim, Se-Young Yun

Sound source separation has attracted attention from Music Information Retrieval(MIR) researchers, since it is related to many MIR tasks such as automatic lyric transcription, singer identification, and voice conversion. In this paper, we propose an intuitive spectrogram-based model for source separation by adapting U-Net. We call it Spectrogram-Channels U-Net, which means each channel of the output corresponds to the spectrogram of separated source itself. The proposed model can be used for not only singing voice separation but also multi-instrument separation by changing only the number of output channels. In addition, we propose a loss function that balances volumes between different sources. Finally, we yield performance that is state-of-the-art on both separation tasks.

📄 PDF Abstract BibTeX arXiv:1810.11520

Code (1)

lucas-dunker/stem-separator-amt pytorch

Tasks

Information RetrievalMusic Information RetrievalRetrievalSinger IdentificationVoice Conversion

Methods 이 논문이 사용한 방법론

Concatenated Skip Connection A Concatenated Skip Connection is a type of skip connection that seeks to reuse features by concatenating them to new layers, allowing more information to be retained from…
ReLU How Do I Communicate to Expedia? How Do I Communicate to Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Live Support & Special Travel…
Max Pooling Max Pooling is a pooling operation that calculates the maximum value for patches of a feature map, and uses it to create a downsampled (pooled) feature map. It is usually…
Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…
U-Net 설명 없음

Similar Papers 제목 키워드 기반

Deep Transform: Cocktail Party Source Separation via Complex Convolution in a Deep Neural Network

2015-04-12 · Andrew J. R. Simpson

Convolutional deep neural networks (DNN) are state of the art in many engineering problems but have not yet addressed the issue of how to deal with complex spectrograms. Here, we use circular statistics to provide a conv…

Semi-blind source separation with multichannel variational autoencoder

2018-08-02 · Hirokazu Kameoka, Li Li, Shota Inoue, Shoji Makino

This paper proposes a multichannel source separation technique called the multichannel variational autoencoder (MVAE) method, which uses a conditional VAE (CVAE) to model and estimate the power spectrograms of the source…

blind source separationDecoder

End-to-end music source separation: is it possible in the waveform domain?

2018-10-29 · Francesc Lluís, Jordi Pons, Xavier Serra

Most of the currently successful source separation techniques use the magnitude spectrogram as input, and are therefore by default omitting part of the signal: the phase. To avoid omitting potentially useful information,…

Deep LearningMusic Source Separation

Hybrid Y-Net Architecture for Singing Voice Separation

2023-03-05 · Rashen Fernando, Pamudu Ranasinghe, Udula Ranasinghe, Janaka Wijayakulasooriya 외

This research paper presents a novel deep learning-based neural network architecture, named Y-Net, for achieving music source separation. The proposed architecture performs end-to-end hybrid source separation by extracti…

Music Source Separation

Hybrid Spectrogram and Waveform Source Separation

2021-11-05 · Alexandre Défossez

Source separation models either work on the spectrogram or waveform domain. In this work, we show how to perform end-to-end hybrid source separation, letting the model decide which domain is best suited for each source, …

Music Source Separation