paper-with-me

홈 › Papers

A Two-Stage U-Net for High-Fidelity Denoising of Historical Recordings

2022-02-17 · Eloi Moliner, Vesa Välimäki

Enhancing the sound quality of historical music recordings is a long-standing problem. This paper presents a novel denoising method based on a fully-convolutional deep neural network. A two-stage U-Net model architecture is designed to model and suppress the degradations with high fidelity. The method processes the time-frequency representation of audio, and is trained using realistic noisy data to jointly remove hiss, clicks, thumps, and other common additive disturbances from old analog discs. The proposed model outperforms previous methods in both objective and subjective metrics. The results of a formal blind listening test show that real gramophone recordings denoised with this method have significantly better quality than the baseline methods. This study shows the importance of realistic training data and the power of deep learning in audio restoration.

📄 PDF Abstract BibTeX arXiv:2202.08702

Code (1)

eloimoliner/denoising-historical-recordings 공식 구현 tf

Tasks

Denoising

Methods 이 논문이 사용한 방법론

ReLU How Do I Communicate to Expedia? How Do I Communicate to Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Live Support & Special Travel…
Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…
Max Pooling Max Pooling is a pooling operation that calculates the maximum value for patches of a feature map, and uses it to create a downsampled (pooled) feature map. It is usually…
Concatenated Skip Connection A Concatenated Skip Connection is a type of skip connection that seeks to reuse features by concatenating them to new layers, allowing more information to be retained from…
U-Net 설명 없음

Similar Papers 제목 키워드 기반

VoiceFixer: A Unified Framework for High-Fidelity Speech Restoration

2022-04-12 · Haohe Liu, Xubo Liu, Qiuqiang Kong, Qiao Tian 외

Speech restoration aims to remove distortions in speech signals. Prior methods mainly focus on a single type of distortion, such as speech denoising or dereverberation. However, speech signals can be degraded by several …

Speech DenoisingSpeech EnhancementVocal Bursts Intensity Prediction

ADNAC: Audio Denoiser using Neural Audio Codec

2025-11-03 · Daniel Jimon, Mircea Vaida, Adriana Stan arxiv

Audio denoising is critical in signal processing, enhancing intelligibility and fidelity for applications like restoring musical recordings. This paper presents a proof-of-concept for adapting a state-of-the-art neural a…

Audio Denoising

DiffGAN-TTS: High-Fidelity and Efficient Text-to-Speech with Denoising Diffusion GANs

2022-01-28 · Songxiang Liu, Dan Su, Dong Yu

Denoising diffusion probabilistic models (DDPMs) are expressive generative models that have been used to solve a variety of speech synthesis problems. However, because of their high sampling costs, DDPMs are difficult to…

DenoisingSpeech Synthesistext-to-speechText to Speech

BEHM-GAN: Bandwidth Extension of Historical Music using Generative Adversarial Networks

2022-04-13 · Eloi Moliner, Vesa Välimäki

Audio bandwidth extension aims to expand the spectrum of narrow-band audio signals. Although this topic has been broadly studied during recent years, the particular problem of extending the bandwidth of historical music …

Bandwidth ExtensionDenoising

Learn to See Faster: Pushing the Limits of High-Speed Camera with Deep Underexposed Image Denoising

2022-11-29 · Weihao Zhuang, Tristan Hascoet, Ryoichi Takashima, Tetsuya Takiguchi

The ability to record high-fidelity videos at high acquisition rates is central to the study of fast moving phenomena. The difficulty of imaging fast moving scenes lies in a trade-off between motion blur and underexposur…

DenoisingImage DenoisingSpecificity