paper-with-me

Papers

UnDiff: Unsupervised Voice Restoration with Unconditional Diffusion Model

2023-06-01 · Anastasiia Iashchenko, Pavel Andreev, Ivan Shchekotov, Nicholas Babaev, Dmitry Vetrov

This paper introduces UnDiff, a diffusion probabilistic model capable of solving various speech inverse tasks. Being once trained for speech waveform generation in an unconditional manner, it can be adapted to different tasks including degradation inversion, neural vocoding, and source separation. In this paper, we, first, tackle the challenging problem of unconditional waveform generation by comparing different neural architectures and preconditioning domains. After that, we demonstrate how the trained unconditional diffusion could be adapted to different tasks of speech processing by the means of recent developments in post-training conditioning of diffusion models. Finally, we demonstrate the performance of the proposed technique on the tasks of bandwidth extension, declipping, vocoding, and speech source separation and compare it to the baselines. The codes are publicly available.

📄 PDF Abstract BibTeX arXiv:2306.00721

Code (0)

등록된 구현이 없습니다.

Tasks

Bandwidth Extensionmodel

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

DDRM-PR: Fourier Phase Retrieval using Denoising Diffusion Restoration Models

2025-01-06 · Mehmet Onurcan Kaya, Figen S. Oktem

Diffusion models have demonstrated their utility as learned priors for solving various inverse problems. However, most existing approaches are limited to linear inverse problems. This paper exploits the efficient and uns…

DenoisingRetrieval

Cold Diffusion for Speech Enhancement

2022-11-04 · Hao Yen, François G. Germain, Gordon Wichern, Jonathan Le Roux

Diffusion models have recently shown promising results for difficult enhancement tasks such as the conditional and unconditional restoration of natural images and audio signals. In this work, we explore the possibility o…

Speech Enhancement

Posterior Transition Modeling for Unsupervised Diffusion-Based Speech Enhancement

2025-07-03 · Mostafa Sadeghi, Jean-Eudes Ayilo, Romain Serizel, Xavier Alameda-Pineda arxiv

We explore unsupervised speech enhancement using diffusion models as expressive generative priors for clean speech. Existing approaches guide the reverse diffusion process using noisy speech through an approximate, noise…

Speech Enhancement

Estimation and Restoration of Unknown Nonlinear Distortion using Diffusion

2025-01-10 · Michal Švento, Eloi Moliner, Lauri Juvela, Alec Wright 외

The restoration of nonlinearly distorted audio signals, alongside the identification of the applied memoryless nonlinear operation, is studied. The paper focuses on the difficult but practically important case in which b…

Audio Effects ModelingQuantizationSpeech Enhancement

Self-Rectifying Diffusion Sampling with Perturbed-Attention Guidance

2024-03-26 · Donghoon Ahn, Hyoungwon Cho, Jaewon Min, Wooseok Jang 외

Recent studies have demonstrated that diffusion models are capable of generating high-quality samples, but their quality heavily depends on sampling guidance techniques, such as classifier guidance (CG) and classifier-fr…

DeblurringDenoisingImage GenerationImage Restoration