paper-with-me

Papers

Unsupervised Noise Adaptive Speech Enhancement by Discriminator-Constrained Optimal Transport

2021-11-11 · NeurIPS 2021 12 · Hsin-Yi Lin, Huan-Hsin Tseng, Xugang Lu, Yu Tsao

This paper presents a novel discriminator-constrained optimal transport network (DOTN) that performs unsupervised domain adaptation for speech enhancement (SE), which is an essential regression task in speech processing. The DOTN aims to estimate clean references of noisy speech in a target domain, by exploiting the knowledge available from the source domain. The domain shift between training and testing data has been reported to be an obstacle to learning problems in diverse fields. Although rich literature exists on unsupervised domain adaptation for classification, the methods proposed, especially in regressions, remain scarce and often depend on additional information regarding the input data. The proposed DOTN approach tactically fuses the optimal transport (OT) theory from mathematical analysis with generative adversarial frameworks, to help evaluate continuous labels in the target domain. The experimental results on two SE tasks demonstrate that by extending the classical OT formulation, our proposed DOTN outperforms previous adversarial domain adaptation frameworks in a purely unsupervised manner.

📄 PDF Abstract BibTeX arXiv:2111.06316

Code (1)

hsinyilin19/discriminator-constrained-optimal-transport-network 공식 구현 pytorch

Tasks

Domain AdaptationSpeech EnhancementUnsupervised Domain Adaptation

Similar Papers 제목 키워드 기반

Enhancing Unsupervised Speech Recognition with Diffusion GANs

2023-03-23 · Xianchao Wu

We enhance the vanilla adversarial training method for unsupervised Automatic Speech Recognition (ASR) by a diffusion-GAN. Our model (1) injects instance noises of various intensities to the generator's output and unlabe…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition+1

Improving Speech Recognition on Noisy Speech via Speech Enhancement with Multi-Discriminators CycleGAN

2021-12-12 · Chia-Yu Li, Ngoc Thang Vu

This paper presents our latest investigations on improving automatic speech recognition for noisy speech via speech enhancement. We propose a novel method named Multi-discriminators CycleGAN to reduce noise of input spee…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Speech Enhancementspeech-recognition+1

Unsupervised Speech Enhancement using Dynamical Variational Auto-Encoders

2021-06-23 · Xiaoyu Bie, Simon Leglaive, Xavier Alameda-Pineda, Laurent Girin

Dynamical variational autoencoders (DVAEs) are a class of deep generative models with latent variables, dedicated to model time series of high-dimensional data. DVAEs can be considered as extensions of the variational au…

Representation LearningSpeech EnhancementTime SeriesTime Series Analysis

Unsupervised speech enhancement with deep dynamical generative speech and noise models

2023-06-13 · Xiaoyu Lin, Simon Leglaive, Laurent Girin, Xavier Alameda-Pineda

This work builds on a previous work on unsupervised speech enhancement using a dynamical variational autoencoder (DVAE) as the clean speech model and non-negative matrix factorization (NMF) as the noise model. We propose…

Speech Enhancement

Unsupervised Speech Enhancement using Data-defined Priors

2025-09-26 · Dominik Klement, Matthew Maciejewski, Sanjeev Khudanpur, Jan Černocký 외 arxiv

The majority of deep learning-based speech enhancement methods require paired clean-noisy speech data. Collecting such data at scale in real-world conditions is infeasible, which has led the community to rely on syntheti…

Speech Enhancement