paper-with-me

홈 › Papers

A Perceptually-Motivated Approach for Low-Complexity, Real-Time Enhancement of Fullband Speech

2020-08-27 · Interspeech 2020 8

Over the past few years, speech enhancement methods based on deep learning have greatly surpassed traditional methods based on spectral subtraction and spectral estimation. Many of these new techniques operate directly in the the short-time Fourier transform (STFT) domain, resulting in a high computational complexity. In this work, we propose PercepNet, an efficient approach that relies on human perception of speech by focusing on the spectral envelope and on the periodicity of the speech. We demonstrate high-quality, real-time enhancement of fullband (48 kHz) speech with less than 5% of a CPU core.

📄 PDF Abstract BibTeX arXiv:2008.04259

Code (2)

cookcodes/percepnet
jzi040941/PercepNet pytorch

Tasks

CPUSpeech Enhancement

Similar Papers 제목 키워드 기반

Personalized PercepNet: Real-time, Low-complexity Target Voice Separation and Enhancement

2021-06-08 · Ritwik Giri, Shrikant Venkataramani, Jean-Marc Valin, Umut Isik 외

The presence of multiple talkers in the surrounding environment poses a difficult challenge for real-time speech communication systems considering the constraints on network size and complexity. In this paper, we present…

DeepFilterNet: Perceptually Motivated Real-Time Speech Enhancement

2023-05-14 · Hendrik Schröter, Tobias Rosenkranz, Alberto N. Escalante-B., Andreas Maier

Multi-frame algorithms for single-channel speech enhancement are able to take advantage from short-time correlations within the speech signal. Deep Filtering (DF) was proposed to directly estimate a complex filter in fre…

CPUSpeech Enhancement

Speech Enhancement with Perceptually-motivated Optimization and Dual Transformations

2022-09-24 · Xucheng Wan, Kai Liu, Ziqing Du, Huan Zhou

To address the monaural speech enhancement problem, numerous research studies have been conducted to enhance speech via operations either in time-domain on the inner-domain learned from the speech mixture or in time--fre…

Speech Enhancement

A two-step backward compatible fullband speech enhancement system

2022-01-26 · Xu Zhang, LianWu Chen, Xiguang Zheng, Xinlei Ren 외

Speech enhancement methods based on deep learning have surpassed traditional methods. While many of these new approaches are operating on the wideband (16kHz) sample rate, a new fullband (48kHz) speech enhancement system…

Speech EnhancementVocal Bursts Valence Prediction

Stable Training of DNN for Speech Enhancement based on Perceptually-Motivated Black-Box Cost Function

2020-02-14 · Masaki Kawanaka, Yuma Koizumi, Ryoichi Miyazaki, Kohei Yatabe

Improving subjective sound quality of enhanced signals is one of the most important missions in speech enhancement. For evaluating the subjective quality, several methods related to perceptually-motivated objective sound…

Reinforcement LearningSpeech Enhancement