paper-with-me

Papers

Investigating Cross-Domain Losses for Speech Enhancement

2020-10-20 · Sherif Abdulatif, Karim Armanious, Jayasankar T. Sajeev, Karim Guirguis, Bin Yang

Recent years have seen a surge in the number of available frameworks for speech enhancement (SE) and recognition. Whether model-based or constructed via deep learning, these frameworks often rely in isolation on either time-domain signals or time-frequency (TF) representations of speech data. In this study, we investigate the advantages of each set of approaches by separately examining their impact on speech intelligibility and quality. Furthermore, we combine the fragmented benefits of time-domain and TF speech representations by introducing two new cross-domain SE frameworks. A quantitative comparative analysis against recent model-based and deep learning SE approaches is performed to illustrate the merit of the proposed frameworks.

📄 PDF Abstract BibTeX arXiv:2010.10468

Code (0)

등록된 구현이 없습니다.

Tasks

Deep LearningSpeech Enhancement

Similar Papers 제목 키워드 기반

Model as Loss: A Self-Consistent Training Paradigm

2025-05-27 · Saisamarth Rajesh Phaye, Milos Cernak, Andrew Harper

Conventional methods for speech enhancement rely on handcrafted loss functions (e.g., time or frequency domain losses) or deep feature losses (e.g., using WavLM or wav2vec), which often fail to capture subtle signal prop…

DecoderSpeech Enhancement

Investigating the effect of residual and highway connections in speech enhancement models

2018-10-22 · NIPS Workshop IRASL 2018 · Anonymous

Residual and skip connections play an important role in many current generative models. Although their theoretical and numerical advantages are understood, their role in speech enhancement systems has not been in…

DenoisingSpeech DenoisingSpeech Enhancement

Unsupervised Speech Enhancement with speech recognition embedding and disentanglement losses

2021-11-16 · Viet Anh Trinh, Sebastian Braun

Speech enhancement has recently achieved great success with various deep learning methods. However, most conventional speech enhancement systems are trained with supervised methods that impose two significant challenges.…

DisentanglementSpeech Enhancementspeech-recognitionSpeech Recognition

Frequency-Weighted Training Losses for Phoneme-Level DNN-based Speech Enhancement

2025-06-23 · Nasser-Eddine Monir, Paul Magron, Romain Serizel

Recent advances in deep learning have significantly improved multichannel speech enhancement algorithms, yet conventional training loss functions such as the scale-invariant signal-to-distortion ratio (SDR) may fail to p…

Speech Enhancement

A consolidated view of loss functions for supervised deep learning-based speech enhancement

2020-09-25

Deep learning-based speech enhancement for real-time applications recently made large advancements. Due to the lack of a tractable perceptual optimization target, many myths around training losses emerged, whereas the co…

Speech Enhancement