paper-with-me

Papers

Single Channel Speech Enhancement Using Outlier Detection

2016-05-04 · Eunjoon Cho, Bowon Lee, Ronald Schafer, Bernard Widrow

Distortion of the underlying speech is a common problem for single-channel speech enhancement algorithms, and hinders such methods from being used more extensively. A dictionary based speech enhancement method that emphasizes preserving the underlying speech is proposed. Spectral patches of clean speech are sampled and clustered to train a dictionary. Given a noisy speech spectral patch, the best matching dictionary entry is selected and used to estimate the noise power at each time-frequency bin. The noise estimation step is formulated as an outlier detection problem, where the noise at each bin is assumed present only if it is an outlier to the corresponding bin of the best matching dictionary entry. This framework assigns higher priority in removing spectral elements that strongly deviate from a typical spoken unit stored in the trained dictionary. Even without the aid of a separate noise model, this method can achieve significant noise reduction for various non-stationary noises, while effectively preserving the underlying speech in more challenging noisy environments.

📄 PDF Abstract BibTeX arXiv:1605.01329

Code (0)

등록된 구현이 없습니다.

Tasks

Noise EstimationOutlier DetectionSpeech Enhancement

Similar Papers 제목 키워드 기반

Multi-channel end-to-end neural network for speech enhancement, source localization, and voice activity detection

2022-06-20 · Yuan Chen, Yicheng Hsu, Mingsian R. Bai

Speech enhancement and source localization has been active research for several decades with a wide range of real-world applications. Recently, the Deep Complex Convolution Recurrent network (DCCRN) has yielded impressiv…

Action DetectionActivity DetectionSpeech Enhancement

Student-Teacher Learning for BLSTM Mask-based Speech Enhancement

2018-03-27

Spectral mask estimation using bidirectional long short-term memory (BLSTM) neural networks has been widely used in various speech enhancement applications, and it has achieved great success when it is applied to multich…

Speech Enhancementspeech-recognitionSpeech Recognition

SRIB-LEAP submission to Far-field Multi-Channel Speech Enhancement Challenge for Video Conferencing

2021-06-24 · R G Prithvi Raj, Rohit Kumar, M K Jayesh, Anurenjan Purushothaman 외

This paper presents the details of the SRIB-LEAP submission to the ConferencingSpeech challenge 2021. The challenge involved the task of multi-channel speech enhancement to improve the quality of far field speech from mi…

Speech Enhancement

Closing the Gap Between Time-Domain Multi-Channel Speech Enhancement on Real and Simulation Conditions

2021-10-27 · Wangyou Zhang, Jing Shi, Chenda Li, Shinji Watanabe 외

The deep learning based time-domain models, e.g. Conv-TasNet, have shown great potential in both single-channel and multi-channel speech enhancement. However, many experiments on the time-domain speech enhancement model …

Speech Enhancementspeech-recognitionSpeech Recognition

Exploring the Potential of Data-Driven Spatial Audio Enhancement Using a Single-Channel Model

2024-04-22 · Arthur N. dos Santos, Bruno S. Masiero, Túlio C. L. Mateus

One key aspect differentiating data-driven single- and multi-channel speech enhancement and dereverberation methods is that both the problem formulation and complexity of the solutions are considerably more challenging i…

Direction of Arrival EstimationSpeech Enhancement