paper-with-me

Papers

Plugin Speech Enhancement: A Universal Speech Enhancement Framework Inspired by Dynamic Neural Network

2024-02-20 · Yanan Chen, Zihao Cui, Yingying Gao, Junlan Feng, Chao Deng, Shilei Zhang

The expectation to deploy a universal neural network for speech enhancement, with the aim of improving noise robustness across diverse speech processing tasks, faces challenges due to the existing lack of awareness within static speech enhancement frameworks regarding the expected speech in downstream modules. These limitations impede the effectiveness of static speech enhancement approaches in achieving optimal performance for a range of speech processing tasks, thereby challenging the notion of universal applicability. The fundamental issue in achieving universal speech enhancement lies in effectively informing the speech enhancement module about the features of downstream modules. In this study, we present a novel weighting prediction approach, which explicitly learns the task relationships from downstream training information to address the core challenge of universal speech enhancement. We found the role of deciding whether to employ data augmentation techniques as crucial downstream training information. This decision significantly impacts the expected speech and the performance of the speech enhancement module. Moreover, we introduce a novel speech enhancement network, the Plugin Speech Enhancement (Plugin-SE). The Plugin-SE is a dynamic neural network that includes the speech enhancement module, gate module, and weight prediction module. Experimental results demonstrate that the proposed Plugin-SE approach is competitive or superior to other joint training methods across various downstream tasks.

📄 PDF Abstract BibTeX arXiv:2402.12746

Code (0)

등록된 구현이 없습니다.

Tasks

Data AugmentationSpeech Enhancement

Similar Papers 제목 키워드 기반

TS-URGENet: A Three-stage Universal Robust and Generalizable Speech Enhancement Network

2025-05-24 · Xiaobin Rong, DaHan Wang, Qinwen Hu, Yushi Wang 외

Universal speech enhancement aims to handle input speech with different distortions and input formats. To tackle this challenge, we present TS-URGENet, a Three-Stage Universal, Robust, and Generalizable speech Enhancemen…

Speech Enhancement

Universal Score-based Speech Enhancement with High Content Preservation

2024-06-18 · Robin Scheibler, Yusuke Fujita, Yuma Shirahata, Tatsuya Komatsu

We propose UNIVERSE++, a universal speech enhancement method based on score-based diffusion and adversarial training. Specifically, we improve the existing UNIVERSE model that decouples clean speech feature extraction an…

Speech Enhancement

Universal Speech Enhancement with Score-based Diffusion

2022-06-07 · Joan Serrà, Santiago Pascual, Jordi Pons, R. Oguz Araz 외

Removing background noise from speech audio has been the subject of considerable effort, especially in recent years due to the rise of virtual communication and amateur recordings. Yet background noise is not the only un…

Speech Enhancement

FINALLY: fast and universal speech enhancement with studio-like quality

2024-10-08 · Nicholas Babaev, Kirill Tamogashev, Azat Saginbaev, Ivan Shchekotov 외

In this paper, we address the challenge of speech enhancement in real-world recordings, which often contain various forms of distortion, such as background noise, reverberation, and microphone artifacts. We revisit the u…

Speech Enhancement

Masked Autoencoders as Universal Speech Enhancer

2026-02-02 · Rajalaxmi Rajagopalan, Ritwik Giri, Zhiqiang Tang, Kyu Han arxiv

Supervised speech enhancement methods have been very successful. However, in practical scenarios, there is a lack of clean speech, and self-supervised learning-based (SSL) speech enhancement methods that offer comparable…

Self-Supervised LearningSpeech Enhancement