paper-with-me

Papers

Phonetic Feedback for Speech Enhancement With and Without Parallel Speech Data

2020-03-03 · Peter Plantinga, Deblin Bagchi, Eric Fosler-Lussier

While deep learning systems have gained significant ground in speech enhancement research, these systems have yet to make use of the full potential of deep learning systems to provide high-level feedback. In particular, phonetic feedback is rare in speech enhancement research even though it includes valuable top-down information. We use the technique of mimic loss to provide phonetic feedback to an off-the-shelf enhancement system, and find gains in objective intelligibility scores on CHiME-4 data. This technique takes a frozen acoustic model trained on clean speech to provide valuable feedback to the enhancement model, even in the case where no parallel speech data is available. Our work is one of the first to show intelligibility improvement for neural enhancement systems without parallel speech data, and we show phonetic feedback can improve a state-of-the-art neural enhancement system trained with parallel speech data.

📄 PDF Abstract BibTeX arXiv:2003.01769

Code (1)

OSU-slatelab/mimic-enhance 공식 구현 pytorch

Tasks

Speech Enhancement

Similar Papers 제목 키워드 기반

Improving Speech Enhancement Performance by Leveraging Contextual Broad Phonetic Class Information

2020-11-15 · Yen-Ju Lu, Chia-Yu Chang, Cheng Yu, Ching-Feng Liu 외

Previous studies have confirmed that by augmenting acoustic features with the place/manner of articulatory features, the speech enhancement (SE) process can be guided to consider the broad phonetic properties of the inpu…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)DenoisingMulti-Task Learning+5

A Systematic Comparison of Phonetic Aware Techniques for Speech Enhancement

2022-06-22 · Or Tal, Moshe Mandel, Felix Kreuk, Yossi Adi

Speech enhancement has seen great improvement in recent years using end-to-end neural networks. However, most models are agnostic to the spoken phonetic content. Recently, several studies suggested phonetic-aware speech …

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Model OptimizationSelf-Supervised Learning+2

Incorporating Broad Phonetic Information for Speech Enhancement

2020-08-13 · Yen-Ju Lu, Chien-Feng Liao, Xugang Lu, Jeih-weih Hung 외

In noisy conditions, knowing speech contents facilitates listeners to more effectively suppress background noise components and to retrieve pure speech signals. Previous studies have also confirmed the benefits of incorp…

DenoisingSpeech Enhancement

Improving Perceptual Quality by Phone-Fortified Perceptual Loss using Wasserstein Distance for Speech Enhancement

2020-10-28 · Tsun-An Hsieh, Cheng Yu, Szu-Wei Fu, Xugang Lu 외

Speech enhancement (SE) aims to improve speech quality and intelligibility, which are both related to a smooth transition in speech segments that may carry linguistic information, e.g. phones and syllables. In this study…

Speech Enhancement

PAAPLoss: A Phonetic-Aligned Acoustic Parameter Loss for Speech Enhancement

2023-02-16 · Muqiao Yang, Joseph Konan, David Bick, Yunyang Zeng 외

Despite rapid advancement in recent years, current speech enhancement models often produce speech that differs in perceptual quality from real clean speech. We propose a learning objective that formalizes differences in …

Speech EnhancementTime SeriesTime Series Analysis