paper-with-me

Papers

A Diffusion-Based Generative Equalizer for Music Restoration

2024-03-27 · Eloi Moliner, Maija Turunen, Filip Elvander, Vesa Välimäki

This paper presents a novel approach to audio restoration, focusing on the enhancement of low-quality music recordings, and in particular historical ones. Building upon a previous algorithm called BABE, or Blind Audio Bandwidth Extension, we introduce BABE-2, which presents a series of improvements. This research broadens the concept of bandwidth extension to \emph{generative equalization}, a novel task that, to the best of our knowledge, has not been explicitly addressed in previous studies. BABE-2 is built around an optimization algorithm utilizing priors from diffusion models, which are trained or fine-tuned using a curated set of high-quality music tracks. The algorithm simultaneously performs two critical tasks: estimation of the filter degradation magnitude response and hallucination of the restored audio. The proposed method is objectively evaluated on historical piano recordings, showing an enhancement over the prior version. The method yields similarly impressive results in rejuvenating the works of renowned vocalists Enrico Caruso and Nellie Melba. This research represents an advancement in the practical restoration of historical music.

📄 PDF Abstract BibTeX arXiv:2403.18636

Code (1)

eloimoliner/babe2 공식 구현 pytorch

Tasks

Bandwidth ExtensionHallucination

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically
Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Latent Fourier Transform

2026-04-20 · Mason Wang, Cheng-Zhi Anna Huang arxiv

We introduce the Latent Fourier Transform (LatentFT), a framework that provides novel frequency-domain controls for generative music models. LatentFT combines a diffusion autoencoder with a latent-space Fourier transform…

VRDMG: Vocal Restoration via Diffusion Posterior Sampling with Multiple Guidance

2023-09-13 · Carlos Hernandez-Olivan, Koichi Saito, Naoki Murata, Chieh-Hsin Lai 외

Restoring degraded music signals is essential to enhance audio quality for downstream music manipulation. Recent diffusion-based music restoration methods have demonstrated impressive performance, and among them, diffusi…

Bandwidth Extension

Token-Based Audio Inpainting via Discrete Diffusion

2025-07-11 · Tali Dror, Iftach Shoham, Moshe Buchris, Oren Gal 외 arxiv

Audio inpainting seeks to restore missing segments in degraded recordings. Previous diffusion-based methods exhibit impaired performance when the missing region is large. We introduce the first approach that applies disc…

Diffusion Models for Audio Restoration

2024-02-15 · Jean-Marie Lemercier, Julius Richter, Simon Welker, Eloi Moliner 외

With the development of audio playback devices and fast data transmission, the demand for high sound quality is rising for both entertainment and communications. In this quest for better sound quality, challenges emerge …

Speech Enhancement

Apollo: Band-sequence Modeling for High-Quality Audio Restoration

2024-09-13 · Kai Li, Yi Luo

Audio restoration has become increasingly significant in modern society, not only due to the demand for high-quality auditory experiences enabled by advanced playback devices, but also because the growing capabilities of…

Computational EfficiencySpeech Enhancement