paper-with-me

홈 › Papers

Stable Training of DNN for Speech Enhancement based on Perceptually-Motivated Black-Box Cost Function

2020-02-14 · Masaki Kawanaka, Yuma Koizumi, Ryoichi Miyazaki, Kohei Yatabe

Improving subjective sound quality of enhanced signals is one of the most important missions in speech enhancement. For evaluating the subjective quality, several methods related to perceptually-motivated objective sound quality assessment (OSQA) have been proposed such as PESQ (perceptual evaluation of speech quality). However, direct use of such measures for training deep neural network (DNN) is not allowed in most cases because popular OSQAs are non-differentiable with respect to DNN parameters. Therefore, the previous study has proposed to approximate the score of OSQAs by an auxiliary DNN so that its gradient can be used for training the primary DNN. One problem with this approach is instability of the training caused by the approximation error of the score. To overcome this problem, we propose to use stabilization techniques borrowed from reinforcement learning. The experiments, aimed to increase the score of PESQ as an example, show that the proposed method (i) can stably train a DNN to increase PESQ, (ii) achieved the state-of-the-art PESQ score on a public dataset, and (iii) resulted in better sound quality than conventional methods based on subjective evaluation.

📄 PDF Abstract BibTeX arXiv:2002.05879

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement LearningSpeech Enhancement

Similar Papers 제목 키워드 기반

A two-step backward compatible fullband speech enhancement system

2022-01-26 · Xu Zhang, LianWu Chen, Xiguang Zheng, Xinlei Ren 외

Speech enhancement methods based on deep learning have surpassed traditional methods. While many of these new approaches are operating on the wideband (16kHz) sample rate, a new fullband (48kHz) speech enhancement system…

Speech EnhancementVocal Bursts Valence Prediction

A Perceptually-Motivated Approach for Low-Complexity, Real-Time Enhancement of Fullband Speech

2020-08-27 · Interspeech 2020 8

Over the past few years, speech enhancement methods based on deep learning have greatly surpassed traditional methods based on spectral subtraction and spectral estimation. Many of these new techniques operate directly i…

CPUSpeech Enhancement

DeepFilterNet: Perceptually Motivated Real-Time Speech Enhancement

2023-05-14 · Hendrik Schröter, Tobias Rosenkranz, Alberto N. Escalante-B., Andreas Maier

Multi-frame algorithms for single-channel speech enhancement are able to take advantage from short-time correlations within the speech signal. Deep Filtering (DF) was proposed to directly estimate a complex filter in fre…

CPUSpeech Enhancement

Speech Enhancement with Perceptually-motivated Optimization and Dual Transformations

2022-09-24 · Xucheng Wan, Kai Liu, Ziqing Du, Huan Zhou

To address the monaural speech enhancement problem, numerous research studies have been conducted to enhance speech via operations either in time-domain on the inner-domain learned from the speech mixture or in time--fre…

Speech Enhancement

Perceptually-motivated Environment-specific Speech Enhancement

2019-05-01 · ICASSP 2019 5 · Jiaqi Su, Adam Finkelstein, Zeyu Jin

This paper introduces a deep learning approach to enhance speech recordings made in a specific environment. A single neural network learns to ameliorate several types of recording artifacts, including noise, reverberatio…

Speech Enhancement