paper-with-me

홈 › Papers

Unsupervised vocal dereverberation with diffusion-based generative models

2022-11-08 · Koichi Saito, Naoki Murata, Toshimitsu Uesaka, Chieh-Hsin Lai, Yuhta Takida, Takao Fukui, Yuki Mitsufuji

Removing reverb from reverberant music is a necessary technique to clean up audio for downstream music manipulations. Reverberation of music contains two categories, natural reverb, and artificial reverb. Artificial reverb has a wider diversity than natural reverb due to its various parameter setups and reverberation types. However, recent supervised dereverberation methods may fail because they rely on sufficiently diverse and numerous pairs of reverberant observations and retrieved data for training in order to be generalizable to unseen observations during inference. To resolve these problems, we propose an unsupervised method that can remove a general kind of artificial reverb for music without requiring pairs of data for training. The proposed method is based on diffusion models, where it initializes the unknown reverberation operator with a conventional signal processing technique and simultaneously refines the estimate with the help of diffusion models. We show through objective and perceptual evaluations that our method outperforms the current leading vocal dereverberation benchmarks.

📄 PDF Abstract BibTeX arXiv:2211.04124

Code (0)

등록된 구현이 없습니다.

Tasks

Diversity

Methods 이 논문이 사용한 방법론

fail 설명 없음
Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

StoRM: A Diffusion-based Stochastic Regeneration Model for Speech Enhancement and Dereverberation

2022-12-22 · Jean-Marie Lemercier, Julius Richter, Simon Welker, Timo Gerkmann

Diffusion models have shown a great ability at bridging the performance gap between predictive and generative approaches for speech enhancement. We have shown that they may even outperform their predictive counterparts f…

Speech DereverberationSpeech Enhancement

Unsupervised Blind Joint Dereverberation and Room Acoustics Estimation with Diffusion Models

2024-08-14 · Jean-Marie Lemercier, Eloi Moliner, Simon Welker, Vesa Välimäki 외

This paper presents an unsupervised method for single-channel blind dereverberation and room impulse response (RIR) estimation, called BUDDy. The algorithm is rooted in Bayesian posterior sampling: it combines a likeliho…

Room Impulse Response (RIR)Speech Dereverberation

BUDDy: Single-Channel Blind Unsupervised Dereverberation with Diffusion Models

2024-05-07 · Eloi Moliner, Jean-Marie Lemercier, Simon Welker, Timo Gerkmann 외

In this paper, we present an unsupervised single-channel method for joint blind dereverberation and room impulse response estimation, based on posterior sampling with diffusion models. We parameterize the reverberation o…

Analysing Diffusion-based Generative Approaches versus Discriminative Approaches for Speech Restoration

2022-11-04 · Jean-Marie Lemercier, Julius Richter, Simon Welker, Timo Gerkmann

Diffusion-based generative models have had a high impact on the computer vision and speech processing communities these past years. Besides data generation tasks, they have also been employed for data restoration tasks l…

Bandwidth ExtensionSpeech DenoisingSpeech DereverberationSpeech Enhancement

GibbsDDRM: A Partially Collapsed Gibbs Sampler for Solving Blind Inverse Problems with Denoising Diffusion Restoration

2023-01-30 · Naoki Murata, Koichi Saito, Chieh-Hsin Lai, Yuhta Takida 외

Pre-trained diffusion models have been successfully used as priors in a variety of linear inverse problems, where the goal is to reconstruct a signal from noisy linear measurements. However, existing approaches require k…

Blind Image DeblurringDeblurringDenoisingImage Deblurring