paper-with-me

Papers

Heterogeneous Target Speech Separation

2022-04-07 · Efthymios Tzinis, Gordon Wichern, Aswin Subramanian, Paris Smaragdis, Jonathan Le Roux

We introduce a new paradigm for single-channel target source separation where the sources of interest can be distinguished using non-mutually exclusive concepts (e.g., loudness, gender, language, spatial location, etc). Our proposed heterogeneous separation framework can seamlessly leverage datasets with large distribution shifts and learn cross-domain representations under a variety of concepts used as conditioning. Our experiments show that training separation models with heterogeneous conditions facilitates the generalization to new concepts with unseen out-of-domain data while also performing substantially higher than single-domain specialist models. Notably, such training leads to more robust learning of new harder source separation discriminative concepts and can yield improvements over permutation invariant training with oracle source selection. We analyze the intrinsic behavior of source separation training with heterogeneous metadata and propose ways to alleviate emerging problems with challenging separation conditions. We release the collection of preparation recipes for all datasets used to further promote research towards this challenging task.

📄 PDF Abstract BibTeX arXiv:2204.03594

Code (0)

등록된 구현이 없습니다.

Tasks

Speech Separation

Similar Papers 제목 키워드 기반

Heterogeneous Separation Consistency Training for Adaptation of Unsupervised Speech Separation

2022-04-23 · Jiangyu Han, Yanhua Long

Recently, supervised speech separation has made great progress. However, limited by the nature of supervised training, most existing separation methods require ground-truth sources and are trained on synthetic datasets. …

Speech Separation

Supervised Speech Separation Based on Deep Learning: An Overview

2017-08-24 · DeLiang Wang, Jitong Chen

Speech separation is the task of separating target speech from background interference. Traditionally, speech separation is studied as a signal processing problem. A more recent approach formulates speech separation as a…

Deep LearningSpeaker SeparationSpeech DereverberationSpeech Enhancement+1

Guided Training: A Simple Method for Single-channel Speaker Separation

2021-03-26 · Hao Li, Xueliang Zhang, Guanglai Gao

Deep learning has shown a great potential for speech separation, especially for speech and non-speech separation. However, it encounters permutation problem for multi-speaker separation where both target and interference…

Speaker SeparationSpeech Separation

Using Optimal Ratio Mask as Training Target for Supervised Speech Separation

2017-09-04 · Shasha Xia, Hao Li, Xueliang Zhang

Supervised speech separation uses supervised learning algorithms to learn a mapping from an input noisy signal to an output target. With the fast development of deep learning, supervised separation has become the most im…

Speech Separation

A Single Speech Enhancement Model Unifying Dereverberation, Denoising, Speaker Counting, Separation, and Extraction

2023-10-12 · Kohei Saijo, Wangyou Zhang, Zhong-Qiu Wang, Shinji Watanabe 외

We propose a multi-task universal speech enhancement (MUSE) model that can perform five speech enhancement (SE) tasks: dereverberation, denoising, speech separation (SS), target speaker extraction (TSE), and speaker coun…

DenoisingSpeech EnhancementSpeech SeparationTarget Speaker Extraction