paper-with-me

홈 › Papers

Recycling an anechoic pre-trained speech separation deep neural network for binaural dereverberation of a single source

2022-08-09 · Sania Gul, Muhammad Salman Khan, Syed Waqar Shah, Ata Ur-Rehman

Reverberation results in reduced intelligibility for both normal and hearing-impaired listeners. This paper presents a novel psychoacoustic approach of dereverberation of a single speech source by recycling a pre-trained binaural anechoic speech separation neural network. As training the deep neural network (DNN) is a lengthy and computationally expensive process, the advantage of using a pre-trained separation network for dereverberation is that the network does not need to be retrained, saving both time and computational resources. The interaural cues of a reverberant source are given to this pretrained neural network to discriminate between the direct path signal and the reverberant speech. The results show an average improvement of 1.3% in signal intelligibility, 0.83 dB in SRMR (signal to reverberation energy ratio) and 0.16 points in perceptual evaluation of speech quality (PESQ) over other state-of-the-art signal processing dereverberation algorithms and 14% in intelligibility and 0.35 points in quality over orthogonal matching pursuit with spectral subtraction (OSS), a machine learning based dereverberation algorithm.

📄 PDF Abstract BibTeX arXiv:2208.04626

Code (0)

등록된 구현이 없습니다.

Tasks

Speech Separation

Similar Papers 제목 키워드 기반

Online Self-Attentive Gated RNNs for Real-Time Speaker Separation

2021-06-25 · Ori Kabeli, Yossi Adi, Zhenyu Tang, Buye Xu 외

Deep neural networks have recently shown great success in the task of blind source separation, both under monaural and binaural settings. Although these methods were shown to produce high-quality separations, they were m…

blind source separationSpeaker Separation

Towards Real-Time Single-Channel Speech Separation in Noisy and Reverberant Environments

2023-03-14 · Julian Neri, Sebastian Braun

Real-time single-channel speech separation aims to unmix an audio stream captured from a single microphone that contains multiple people talking at once, environmental noise, and reverberation into multiple de-reverberat…

DecoderSpeech Separation

Cognitive performance in open-plan office acoustic simulations: Effects of room acoustics and semantics but not spatial separation of sound sources

2023-06-13 · Manuj Yadav, Markus Georgi, Larissa Leist, Maria Klatte 외

The irrelevant sound effect (ISE) characterizes short-term memory performance impairment during irrelevant sounds relative to quiet. Irrelevant sound presentation in most laboratory-based ISE studies has been rather limi…

Online Binaural Speech Separation of Moving Speakers With a Wavesplit Network

2023-03-13 · Cong Han, Nima Mesgarani

Binaural speech separation in real-world scenarios often involves moving speakers. Most current speech separation methods use utterance-level permutation invariant training (u-PIT) for training. In inference time, howeve…

Online ClusteringSpeaker SeparationSpeech Separation

Monaural source separation: From anechoic to reverberant environments

2021-11-15 · Tobias Cord-Landwehr, Christoph Boeddeker, Thilo von Neumann, Catalin Zorila 외

Impressive progress in neural network-based single-channel speech source separation has been made in recent years. But those improvements have been mostly reported on anechoic data, a situation that is hardly met in prac…