paper-with-me

Papers

Speech Signal Improvement Using Causal Generative Diffusion Models

2023-03-15 · Julius Richter, Simon Welker, Jean-Marie Lemercier, Bunlong Lay, Tal Peer, Timo Gerkmann

In this paper, we present a causal speech signal improvement system that is designed to handle different types of distortions. The method is based on a generative diffusion model which has been shown to work well in scenarios with missing data and non-linear corruptions. To guarantee causal processing, we modify the network architecture of our previous work and replace global normalization with causal adaptive gain control. We generate diverse training data containing a broad range of distortions. This work was performed in the context of an "ICASSP Signal Processing Grand Challenge" and submitted to the non-real-time track of the "Speech Signal Improvement Challenge 2023", where it was ranked fifth.

📄 PDF Abstract BibTeX arXiv:2303.08674

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Conditional Diffusion Probabilistic Model for Speech Enhancement

2022-02-10 · Yen-Ju Lu, Zhong-Qiu Wang, Shinji Watanabe, Alexander Richard 외

Speech enhancement is a critical component of many user-oriented audio applications, yet current systems still suffer from distorted and unnatural outputs. While generative models have shown strong potential in speech sy…

modelSpeech EnhancementSpeech Synthesis

A Survey on Audio Diffusion Models: Text To Speech Synthesis and Enhancement in Generative AI

2023-03-23 · Chenshuang Zhang, Chaoning Zhang, Sheng Zheng, Mengchun Zhang 외

Generative AI has demonstrated impressive performance in various fields, among which speech synthesis is an interesting direction. With the diffusion model as the most popular generative model, numerous works have attemp…

Speech EnhancementSpeech SynthesisSurveytext-to-speech+2

LatentSpeech: Latent Diffusion for Text-To-Speech Generation

2024-12-11 · Haowei Lou, Helen Paik, Pari Delir Haghighi, Wen Hu 외

Diffusion-based Generative AI gains significant attention for its superior performance over other generative techniques like Generative Adversarial Networks and Variational Autoencoders. While it has achieved notable adv…

text-to-speechText to Speech

uSee: Unified Speech Enhancement and Editing with Conditional Diffusion Models

2023-10-02 · Muqiao Yang, Chunlei Zhang, Yong Xu, Zhongweiyang Xu 외

Speech enhancement aims to improve the quality of speech signals in terms of quality and intelligibility, and speech editing refers to the process of editing the speech according to specific user needs. In this paper, we…

DenoisingSelf-Supervised LearningSpeech DenoisingSpeech Enhancement

Investigating the Effects of Diffusion-based Conditional Generative Speech Models Used for Speech Enhancement on Dysarthric Speech

2024-12-18 · Joanna Reszka, Parvaneh Janbakhshi, Tilak Purohit, Sadegh Mohammadi

In this study, we aim to explore the effect of pre-trained conditional generative speech models for the first time on dysarthric speech due to Parkinson's disease recorded in an ideal/non-noisy condition. Considering one…

Speech Enhancement