paper-with-me

Papers

Diffusion-based Speech Enhancement with Schrödinger Bridge and Symmetric Noise Schedule

2024-09-08 · Siyi Wang, Siyi Liu, Andrew Harper, Paul Kendrick, Mathieu Salzmann, Milos Cernak

Recently, diffusion-based generative models have demonstrated remarkable performance in speech enhancement tasks. However, these methods still encounter challenges, including the lack of structural information and poor performance in low Signal-to-Noise Ratio (SNR) scenarios. To overcome these challenges, we propose the Schr\"oodinger Bridge-based Speech Enhancement (SBSE) method, which learns the diffusion processes directly between the noisy input and the clean distribution, unlike conventional diffusion-based speech enhancement systems that learn data to Gaussian distributions. To enhance performance in extremely noisy conditions, we introduce a two-stage system incorporating ratio mask information into the diffusion-based generative model. Our experimental results show that our proposed SBSE method outperforms all the baseline models and achieves state-of-the-art performance, especially in low SNR conditions. Importantly, only a few inference steps are required to achieve the best result.

📄 PDF Abstract BibTeX arXiv:2409.05116

Code (0)

등록된 구현이 없습니다.

Tasks

Speech Enhancement

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Robust Speech Recognition with Schrödinger Bridge-Based Speech Enhancement

2025-05-07 · Rauf Nasretdinov, Roman Korostik, Ante Jukić

In this work, we investigate application of generative speech enhancement to improve the robustness of ASR models in noisy and reverberant conditions. We employ a recently-proposed speech enhancement model based on Schr\…

Robust Speech RecognitionSpeech Enhancementspeech-recognitionSpeech Recognition

Investigating Training Objectives for Generative Speech Enhancement

2024-09-16 · Julius Richter, Danilo de Oliveira, Timo Gerkmann

Generative speech enhancement has recently shown promising advancements in improving speech quality in noisy environments. Multiple diffusion-based frameworks exist, each employing distinct training objectives and learni…

Speech Enhancement

Schrödinger Bridge for Generative Speech Enhancement

2024-07-22 · Ante Jukić, Roman Korostik, Jagadeesh Balam, Boris Ginsburg

This paper proposes a generative speech enhancement model based on Schr\"odinger bridge (SB). The proposed model is employing a tractable SB to formulate a data-to-data process between the clean speech distribution and t…

DenoisingSpeech DenoisingSpeech DereverberationSpeech Enhancement

Schrödinger Bridge Mamba for One-Step Speech Enhancement

2025-10-19 · Jing Yang, Sirui Wang, Chao Wu, Lei Guo 외 arxiv

We present Schrödinger Bridge Mamba (SBM), a novel model for efficient speech enhancement by integrating the Schrödinger Bridge (SB) training paradigm and the Mamba architecture. Experiments of joint denoising and dereve…

Speech Enhancement

Regularized Schrödinger Bridge via Distortion-Perception Perturbation for High-Fidelity Speech Enhancement

2025-11-12 · Qing Yao, Lijian Gao, Qirong Mao, Ming Dong arxiv

Speech enhancement (SE) requires high-fidelity reconstruction of clean speech that preserves linguistic and paralinguistic cues while maintaining high perceptual quality. Recently, Schrödinger Bridge (SB), a family of di…

Speech Enhancement