paper-with-me

Papers

Incorporating Broad Phonetic Information for Speech Enhancement

2020-08-13 · Yen-Ju Lu, Chien-Feng Liao, Xugang Lu, Jeih-weih Hung, Yu Tsao

In noisy conditions, knowing speech contents facilitates listeners to more effectively suppress background noise components and to retrieve pure speech signals. Previous studies have also confirmed the benefits of incorporating phonetic information in a speech enhancement (SE) system to achieve better denoising performance. To obtain the phonetic information, we usually prepare a phoneme-based acoustic model, which is trained using speech waveforms and phoneme labels. Despite performing well in normal noisy conditions, when operating in very noisy conditions, however, the recognized phonemes may be erroneous and thus misguide the SE process. To overcome the limitation, this study proposes to incorporate the broad phonetic class (BPC) information into the SE process. We have investigated three criteria to build the BPC, including two knowledge-based criteria: place and manner of articulatory and one data-driven criterion. Moreover, the recognition accuracies of BPCs are much higher than that of phonemes, thus providing more accurate phonetic information to guide the SE process under very noisy conditions. Experimental results demonstrate that the proposed SE with the BPC information framework can achieve notable performance improvements over the baseline system and an SE system using monophonic information in terms of both speech quality intelligibility on the TIMIT dataset.

📄 PDF Abstract BibTeX arXiv:2008.07618

Code (0)

등록된 구현이 없습니다.

Tasks

DenoisingSpeech Enhancement

Similar Papers 제목 키워드 기반

Improving Speech Enhancement Performance by Leveraging Contextual Broad Phonetic Class Information

2020-11-15 · Yen-Ju Lu, Chia-Yu Chang, Cheng Yu, Ching-Feng Liu 외

Previous studies have confirmed that by augmenting acoustic features with the place/manner of articulatory features, the speech enhancement (SE) process can be guided to consider the broad phonetic properties of the inpu…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)DenoisingMulti-Task Learning+5

A Systematic Comparison of Phonetic Aware Techniques for Speech Enhancement

2022-06-22 · Or Tal, Moshe Mandel, Felix Kreuk, Yossi Adi

Speech enhancement has seen great improvement in recent years using end-to-end neural networks. However, most models are agnostic to the spoken phonetic content. Recently, several studies suggested phonetic-aware speech …

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Model OptimizationSelf-Supervised Learning+2

Phonetic Feedback for Speech Enhancement With and Without Parallel Speech Data

2020-03-03 · Peter Plantinga, Deblin Bagchi, Eric Fosler-Lussier

While deep learning systems have gained significant ground in speech enhancement research, these systems have yet to make use of the full potential of deep learning systems to provide high-level feedback. In particular, …

Speech Enhancement

Improving Voice Separation by Incorporating End-to-end Speech Recognition

2019-11-29 · Naoya Takahashi, Mayank Kumar Singh, Sakya Basak, Parthasaarathy Sudarsanam 외

Despite recent advances in voice separation methods, many challenges remain in realistic scenarios such as noisy recording and the limits of available data. In this work, we propose to explicitly incorporate the phonetic…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition+2

Improving Perceptual Quality by Phone-Fortified Perceptual Loss using Wasserstein Distance for Speech Enhancement

2020-10-28 · Tsun-An Hsieh, Cheng Yu, Szu-Wei Fu, Xugang Lu 외

Speech enhancement (SE) aims to improve speech quality and intelligibility, which are both related to a smooth transition in speech segments that may carry linguistic information, e.g. phones and syllables. In this study…

Speech Enhancement