UnDiff: Unsupervised Voice Restoration with Unconditional Diffusion Model
This paper introduces UnDiff, a diffusion probabilistic model capable of solving various speech inverse tasks. Being once trained for speech waveform generation in an unconditional manner, it can be adapted to different tasks including degradation inversion, neural vocoding, and source separation. In this paper, we, first, tackle the challenging problem of unconditional waveform generation by comparing different neural architectures and preconditioning domains. After that, we demonstrate how the trained unconditional diffusion could be adapted to different tasks of speech processing by the means of recent developments in post-training conditioning of diffusion models. Finally, we demonstrate the performance of the proposed technique on the tasks of bandwidth extension, declipping, vocoding, and speech source separation and compare it to the baselines. The codes are publicly available.
Code (0)
등록된 구현이 없습니다.
Tasks
Bandwidth ExtensionmodelMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
DDRM-PR: Fourier Phase Retrieval using Denoising Diffusion Restoration Models
Diffusion models have demonstrated their utility as learned priors for solving various inverse problems. However, most existing approaches are limited to linear inverse problems. This paper exploits the efficient and uns…
DenoisingRetrievalCold Diffusion for Speech Enhancement
Diffusion models have recently shown promising results for difficult enhancement tasks such as the conditional and unconditional restoration of natural images and audio signals. In this work, we explore the possibility o…
Speech EnhancementPosterior Transition Modeling for Unsupervised Diffusion-Based Speech Enhancement
We explore unsupervised speech enhancement using diffusion models as expressive generative priors for clean speech. Existing approaches guide the reverse diffusion process using noisy speech through an approximate, noise…
Speech EnhancementEstimation and Restoration of Unknown Nonlinear Distortion using Diffusion
The restoration of nonlinearly distorted audio signals, alongside the identification of the applied memoryless nonlinear operation, is studied. The paper focuses on the difficult but practically important case in which b…
Audio Effects ModelingQuantizationSpeech EnhancementSelf-Rectifying Diffusion Sampling with Perturbed-Attention Guidance
Recent studies have demonstrated that diffusion models are capable of generating high-quality samples, but their quality heavily depends on sampling guidance techniques, such as classifier guidance (CG) and classifier-fr…
DeblurringDenoisingImage GenerationImage Restoration