paper-with-me

홈 › Papers

Active Restoration of Lost Audio Signals Using Machine Learning and Latent Information

2021-11-21 · Zohra Adila Cheddad, Abbas Cheddad

Digital audio signal reconstruction of a lost or corrupt segment using deep learning algorithms has been explored intensively in recent years. Nevertheless, prior traditional methods with linear interpolation, phase coding and tone insertion techniques are still in vogue. However, we found no research work on reconstructing audio signals with the fusion of dithering, steganography, and machine learning regressors. Therefore, this paper proposes the combination of steganography, halftoning (dithering), and state-of-the-art shallow and deep learning methods. The results (including comparing the SPAIN, Autoregressive, deep learning-based, graph-based, and other methods) are evaluated with three different metrics. The observations from the results show that the proposed solution is effective and can enhance the reconstruction of audio signals performed by the side information (e.g., Latent representation) steganography provides. Moreover, this paper proposes a novel framework for reconstruction from heavily compressed embedded audio data using halftoning (i.e., dithering) and machine learning, which we termed the HCR (halftone-based compression and reconstruction). This work may trigger interest in optimising this approach and/or transferring it to different domains (i.e., image reconstruction). Compared to existing methods, we show improvement in the inpainting performance in terms of signal-to-noise ratio (SNR), the objective difference grade (ODG) and Hansen's audio quality metric. In particular, our proposed framework outperformed the learning-based methods (D2WGAN and SG) and the traditional statistical algorithms (e.g., SPAIN, TDC, WCP).

📄 PDF Abstract BibTeX arXiv:2111.10891

Code (0)

등록된 구현이 없습니다.

Tasks

Audio inpaintingDeep LearningImage Reconstruction

Methods 이 논문이 사용한 방법론

Inpainting Train a convolutional neural network to generate the contents of an arbitrary image region conditioned on its surroundings.

Similar Papers 제목 키워드 기반

Blind Restoration of Real-World Audio by 1D Operational GANs

2022-12-30 · Turker Ince, Serkan Kiranyaz, Ozer Can Devecioglu, Muhammad Salman Khan 외

Objective: Despite numerous studies proposed for audio restoration in the literature, most of them focus on an isolated restoration problem such as denoising or dereverberation, ignoring other artifacts. Moreover, assumi…

Denoising

Diffusion Models for Audio Restoration

2024-02-15 · Jean-Marie Lemercier, Julius Richter, Simon Welker, Eloi Moliner 외

With the development of audio playback devices and fast data transmission, the demand for high sound quality is rising for both entertainment and communications. In this quest for better sound quality, challenges emerge …

Speech Enhancement

Inpainting of long audio segments with similarity graphs

2016-07-22 · Nathanael Perraudin, Nicki Holighaus, Piotr Majdak, Peter Balazs

We present a novel method for the compensation of long duration data loss in audio signals, in particular music. The concealment of such signal defects is based on a graph that encodes signal structure in terms of time-p…

On the Design of Deep Priors for Unsupervised Audio Restoration

2021-04-14 · Vivek Sivaraman Narayanaswamy, Jayaraman J. Thiagarajan, Andreas Spanias

Unsupervised deep learning methods for solving audio restoration problems extensively rely on carefully tailored neural architectures that carry strong inductive biases for defining priors in the time or spectral domain.…

Audio DenoisingDenoising

VRDMG: Vocal Restoration via Diffusion Posterior Sampling with Multiple Guidance

2023-09-13 · Carlos Hernandez-Olivan, Koichi Saito, Naoki Murata, Chieh-Hsin Lai 외

Restoring degraded music signals is essential to enhance audio quality for downstream music manipulation. Recent diffusion-based music restoration methods have demonstrated impressive performance, and among them, diffusi…

Bandwidth Extension