paper-with-me

홈 › Papers

Surrogate Source Model Learning for Determined Source Separation

2020-11-11 · Robin Scheibler, Masahito Togami

We propose to learn surrogate functions of universal speech priors for determined blind speech separation. Deep speech priors are highly desirable due to their high modelling power, but are not compatible with state-of-the-art independent vector analysis based on majorization-minimization (AuxIVA), since deriving the required surrogate function is not easy, nor always possible. Instead, we do away with exact majorization and directly approximate the surrogate. Taking advantage of iterative source steering (ISS) updates, we back propagate the permutation invariant separation loss through multiple iterations of AuxIVA. ISS lends itself well to this task due to its lower complexity and lack of matrix inversion. Experiments show large improvements in terms of scale invariant signal-to-distortion (SDR) ratio and word error rate compared to baseline methods. Training is done on two speakers mixtures and we experiment with two losses, SDR and coherence. We find that the learnt approximate surrogate generalizes well on mixtures of three and four speakers without any modification. We also demonstrate generalization to a different variation of the AuxIVA update equations. The SDR loss leads to fastest convergence in iterations, while coherence leads to the lowest word error rate (WER). We obtain as much as 36 % reduction in WER.

📄 PDF Abstract BibTeX arXiv:2011.05540

Code (0)

등록된 구현이 없습니다.

Tasks

modelSpeech Separation

Similar Papers 제목 키워드 기반

Generalized Multichannel Variational Autoencoder for Underdetermined Source Separation

2018-09-29 · Shogo Seki, Hirokazu Kameoka, Li Li, Tomoki Toda 외

This paper deals with a multichannel audio source separation problem under underdetermined conditions. Multichannel Non-negative Matrix Factorization (MNMF) is one of powerful approaches, which adopts the NMF concept for…

Audio Source Separation

Bayesian Non-Parametric Multi-Source Modelling Based Determined Blind Source Separation

2019-04-08 · Chaitanya Narisetty, Tatsuya Komatsu, Reishi Kondo

This paper proposes a determined blind source separation method using Bayesian non-parametric modelling of sources. Conventionally source signals are separated from a given set of mixture signals by modelling them using …

blind source separation

Phase Unmixing : Multichannel Source Separation with Magnitude Constraints

2016-09-30 · Antoine Deleforge, Yann Traonmilin

We consider the problem of estimating the phases of K mixed complex signals from a multichannel observation, when the mixing matrix and signal magnitudes are known. This problem can be cast as a non-convex quadratically …

On Ambisonic Source Separation with Spatially Informed Non-negative Tensor Factorization

2025-01-17 · Mateusz Guzik, Konrad Kowalczyk

This article presents a Non-negative Tensor Factorization based method for sound source separation from Ambisonic microphone signals. The proposed method enables the use of prior knowledge about the Directions-of-Arrival…

Bootstrapping single-channel source separation via unsupervised spatial clustering on stereo mixtures

2018-11-06 · Prem Seetharaman, Gordon Wichern, Jonathan Le Roux, Bryan Pardo

Separating an audio scene into isolated sources is a fundamental problem in computer audition, analogous to image segmentation in visual scene analysis. Source separation systems based on deep learning are currently the …

ClusteringImage SegmentationSemantic SegmentationUnsupervised Spatial Clustering