paper-with-me

Papers

Time-domain Speech Enhancement with Generative Adversarial Learning

2021-03-30 · Feiyang Xiao, Jian Guan, Qiuqiang Kong, Wenwu Wang

Speech enhancement aims to obtain speech signals with high intelligibility and quality from noisy speech. Recent work has demonstrated the excellent performance of time-domain deep learning methods, such as Conv-TasNet. However, these methods can be degraded by the arbitrary scales of the waveform induced by the scale-invariant signal-to-noise ratio (SI-SNR) loss. This paper proposes a new framework called Time-domain Speech Enhancement Generative Adversarial Network (TSEGAN), which is an extension of the generative adversarial network (GAN) in time-domain with metric evaluation to mitigate the scaling problem, and provide model training stability, thus achieving performance improvement. In addition, we provide a new method based on objective function mapping for the theoretical analysis of the performance of Metric GAN, and explain why it is better than the Wasserstein GAN. Experiments conducted demonstrate the effectiveness of our proposed method, and illustrate the advantage of Metric GAN.

📄 PDF Abstract BibTeX arXiv:2103.16149

Code (1)

littleflyingsheep/tsegan 공식 구현 pytorch

Tasks

Generative Adversarial NetworkSpeech Enhancement

Similar Papers 제목 키워드 기반

On the Use of Audio Fingerprinting Features for Speech Enhancement with Generative Adversarial Network

2020-07-27 · Farnood Faraji, Yazid Attabi, Benoit Champagne, Wei-Ping Zhu

The advent of learning-based methods in speech enhancement has revived the need for robust and reliable training features that can compactly represent speech signals while preserving their vital information. Time-frequen…

Generative Adversarial NetworkSpeech Enhancement

A Flow-Based Neural Network for Time Domain Speech Enhancement

2021-06-16 · Martin Strauss, Bernd Edler

Speech enhancement involves the distinction of a target speech signal from an intrusive background. Although generative approaches using Variational Autoencoders or Generative Adversarial Networks (GANs) have increasingl…

Density EstimationSpeech EnhancementSpeech Synthesis

Towards Generalized Speech Enhancement with Generative Adversarial Networks

2019-04-06 · Santiago Pascual, Joan Serrà, Antonio Bonafonte

The speech enhancement task usually consists of removing additive noise or reverberation that partially mask spoken utterances, affecting their intelligibility. However, little attention is drawn to other, perhaps more a…

Generative Adversarial NetworkSpeech Enhancement

FNSE-SBGAN: Far-field Speech Enhancement with Schrodinger Bridge and Generative Adversarial Networks

2025-03-17 · Tong Lei, Qinwen Hu, Ziyao Lin, Andong Li 외

The prevailing method for neural speech enhancement predominantly utilizes fully-supervised deep learning with simulated pairs of far-field noisy-reverberant speech and clean speech. Nonetheless, these models frequently …

Speech Enhancement

SEGAN: Speech Enhancement Generative Adversarial Network

2017-03-28 · Santiago Pascual, Antonio Bonafonte, Joan Serrà

Current speech enhancement techniques operate on the spectral domain and/or exploit some higher-level feature. The majority of them tackle a limited number of noise conditions and rely on first-order statistics. To circu…

Generative Adversarial NetworkSpeech Enhancement