paper-with-me

Papers

Unsupervised Interpretable Representation Learning for Singing Voice Separation

2020-03-03 · Stylianos I. Mimilakis, Konstantinos Drossos, Gerald Schuller

In this work, we present a method for learning interpretable music signal representations directly from waveform signals. Our method can be trained using unsupervised objectives and relies on the denoising auto-encoder model that uses a simple sinusoidal model as decoding functions to reconstruct the singing voice. To demonstrate the benefits of our method, we employ the obtained representations to the task of informed singing voice separation via binary masking, and measure the obtained separation quality by means of scale-invariant signal to distortion ratio. Our findings suggest that our method is capable of learning meaningful representations for singing voice separation, while preserving conveniences of the the short-time Fourier transform like non-negativity, smoothness, and reconstruction subject to time-frequency masking, that are desired in audio and music source separation.

📄 PDF Abstract BibTeX arXiv:2003.01567

Code (1)

Js-Mim/rl_singing_voice 공식 구현 pytorch

Tasks

DenoisingMusic Source SeparationRepresentation Learning

Similar Papers 제목 키워드 기반

A fully differentiable model for unsupervised singing voice separation

2024-01-30 · Gael Richard, Pierre Chouteau, Bernardo Torres

A novel model was recently proposed by Schulze-Forster et al. in [1] for unsupervised music source separation. This model allows to tackle some of the major shortcomings of existing source separation frameworks. Specific…

Music Source Separation

MedleyVox: An Evaluation Dataset for Multiple Singing Voices Separation

2022-11-14 · Chang-Bin Jeon, Hyeongi Moon, Keunwoo Choi, Ben Sangbae Chon 외

Separation of multiple singing voices into each voice is a rarely studied area in music source separation research. The absence of a benchmark dataset has hindered its progress. In this paper, we present an evaluation da…

Music Source SeparationSuper-Resolution

Informed Group-Sparse Representation for Singing Voice Separation

2018-01-09 · Tak-Shing T. Chan, Yi-Hsuan Yang

Singing voice separation attempts to separate the vocal and instrumental parts of a music recording, which is a fundamental problem in music information retrieval. Recent work on singing voice separation has shown that t…

Information RetrievalMusic Information RetrievalRetrieval

A cappella: Audio-visual Singing Voice Separation

2021-04-20 · Juan F. Montesinos, Venkatesh S. Kadandale, Gloria Haro

The task of isolating a target singing voice in music videos has useful applications. In this work, we explore the single-channel singing voice separation problem from a multimodal perspective, by jointly learning from a…

Music Source SeparationSpeech Separation

PitchNet: Unsupervised Singing Voice Conversion with Pitch Adversarial Network

2019-12-04 · Chengqi Deng, Chengzhu Yu, Heng Lu, Chao Weng 외

Singing voice conversion is to convert a singer's voice to another one's voice without changing singing content. Recent work shows that unsupervised singing voice conversion can be achieved with an autoencoder-based appr…

DecoderMusic GenerationTranslationVoice Conversion