paper-with-me

Papers

Semi-Supervised Monaural Singing Voice Separation With a Masking Network Trained on Synthetic Mixtures

2018-12-14 · Michael Michelashvili, Sagie Benaim, Lior Wolf

We study the problem of semi-supervised singing voice separation, in which the training data contains a set of samples of mixed music (singing and instrumental) and an unmatched set of instrumental music. Our solution employs a single mapping function g, which, applied to a mixed sample, recovers the underlying instrumental music, and, applied to an instrumental sample, returns the same sample. The network g is trained using purely instrumental samples, as well as on synthetic mixed samples that are created by mixing reconstructed singing voices with random instrumental samples. Our results indicate that we are on a par with or better than fully supervised methods, which are also provided with training samples of unmixed singing voices, and are better than other recent semi-supervised methods.

📄 PDF Abstract BibTeX arXiv:1812.06087

Code (1)

sagiebenaim/Singing pytorch

Tasks

Music Source SeparationSpeech Separation

Similar Papers 제목 키워드 기반

Joint Optimization of Masks and Deep Recurrent Neural Networks for Monaural Source Separation

2015-02-13 · Po-Sen Huang, Minje Kim, Mark Hasegawa-Johnson, Paris Smaragdis

Monaural source separation is important for many real world applications. It is challenging because, with only a single channel of information available, without any constraints, an infinite number of solutions are possi…

DenoisingSpeech DenoisingSpeech Separation

HTMD-Net: A Hybrid Masking-Denoising Approach to Time-Domain Monaural Singing Voice Separation

2021-03-07 · Christos Garoufis, Athanasia Zlatintsi, Petros Maragos

The advent of deep learning has led to the prevalence of deep neural network architectures for monaural music source separation, with end-to-end approaches that operate directly on the waveform level increasingly receivi…

Computational EfficiencyDenoisingMusic Source Separation

Evolving Multi-Resolution Pooling CNN for Monaural Singing Voice Separation

2020-08-03 · Weitao Yuan, Bofei Dong, Shengbei Wang, Masashi Unoki 외

Monaural Singing Voice Separation (MSVS) is a challenging task and has been studied for decades. Deep neural networks (DNNs) are the current state-of-the-art methods for MSVS. However, the existing DNNs are often designe…

Neural Architecture Search

Improved singing voice separation with chromagram-based pitch-aware remixing

2022-03-28 · Siyuan Yuan, Zhepei Wang, Umut Isik, Ritwik Giri 외

Singing voice separation aims to separate music into vocals and accompaniment components. One of the major constraints for the task is the limited amount of training data with separated vocals. Data augmentation techniqu…

Data Augmentation

Unsupervised Interpretable Representation Learning for Singing Voice Separation

2020-03-03 · Stylianos I. Mimilakis, Konstantinos Drossos, Gerald Schuller

In this work, we present a method for learning interpretable music signal representations directly from waveform signals. Our method can be trained using unsupervised objectives and relies on the denoising auto-encoder m…

DenoisingMusic Source SeparationRepresentation Learning