paper-with-me

Papers

A Cold Diffusion Approach for Percussive Dereverberation

2026-05-11 · Dimos Makris, András Barják, Maximos Kaliakatsos-Papakostas arxiv

Most recent advances in audio dereverberation focus almost exclusively on speech, leaving percussive and drum signals largely unexplored despite their importance in music production. Percussive dereverberation poses distinct challenges due to sharp transients and dense temporal structure. In this work, we propose a cold diffusion framework for dereverberating stereo drum stems (downmixes), modeling reverberation as a deterministic degradation process that progressively transforms anechoic signals into reverberant ones. We investigate two reverse-process parameterizations, Direct (next-state) and a Delta-normalized residual (velocity-style) prediction, and implement the framework using both a UNet and a diffusion Transformer backbone. The models are trained and evaluated on curated datasets comprising both acoustic and electronic drum recordings, with reverberation generated using a combination of synthetic and real room impulse responses. Extensive experiments on in-domain and fully out-of-domain test sets demonstrate that the proposed method consistently outperforms strong score-based and conditional diffusion baselines, evaluated using signal-based and perceptual metrics tailored to percussive audio.

📄 PDF Abstract BibTeX arXiv:2605.10256

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Diffusion Posterior Sampling for Informed Single-Channel Dereverberation

2023-06-21 · Jean-Marie Lemercier, Simon Welker, Timo Gerkmann

We present in this paper an informed single-channel dereverberation method based on conditional generation with diffusion models. With knowledge of the room impulse response, the anechoic utterance is generated via rever…

Unsupervised Blind Joint Dereverberation and Room Acoustics Estimation with Diffusion Models

2024-08-14 · Jean-Marie Lemercier, Eloi Moliner, Simon Welker, Vesa Välimäki 외

This paper presents an unsupervised method for single-channel blind dereverberation and room impulse response (RIR) estimation, called BUDDy. The algorithm is rooted in Bayesian posterior sampling: it combines a likeliho…

Room Impulse Response (RIR)Speech Dereverberation

BUDDy: Single-Channel Blind Unsupervised Dereverberation with Diffusion Models

2024-05-07 · Eloi Moliner, Jean-Marie Lemercier, Simon Welker, Timo Gerkmann 외

In this paper, we present an unsupervised single-channel method for joint blind dereverberation and room impulse response estimation, based on posterior sampling with diffusion models. We parameterize the reverberation o…

Analysing Diffusion-based Generative Approaches versus Discriminative Approaches for Speech Restoration

2022-11-04 · Jean-Marie Lemercier, Julius Richter, Simon Welker, Timo Gerkmann

Diffusion-based generative models have had a high impact on the computer vision and speech processing communities these past years. Besides data generation tasks, they have also been employed for data restoration tasks l…

Bandwidth ExtensionSpeech DenoisingSpeech DereverberationSpeech Enhancement

Unsupervised vocal dereverberation with diffusion-based generative models

2022-11-08 · Koichi Saito, Naoki Murata, Toshimitsu Uesaka, Chieh-Hsin Lai 외

Removing reverb from reverberant music is a necessary technique to clean up audio for downstream music manipulations. Reverberation of music contains two categories, natural reverb, and artificial reverb. Artificial reve…

Diversity