paper-with-me

홈 › Papers

Audio inpainting of music by means of neural networks

2018-10-29 · Andrés Marafioti, Nicki Holighaus, Piotr Majdak, Nathanaël Perraudin

We studied the ability of deep neural networks (DNNs) to restore missing audio content based on its context, a process usually referred to as audio inpainting. We focused on gaps in the range of tens of milliseconds. The proposed DNN structure was trained on audio signals containing music and musical instruments, separately, with 64-ms long gaps. The input to the DNN was the context, i.e., the signal surrounding the gap, transformed into time-frequency (TF) coefficients. Our results were compared to those obtained from a reference method based on linear predictive coding (LPC). For music, our DNN significantly outperformed the reference method, demonstrating a generally good usability of the proposed DNN structure for inpainting complex audio signals like music.

📄 PDF Abstract BibTeX arXiv:1810.12138

Code (1)

andimarafioti/audioContextEncoder tf

Tasks

Audio GenerationAudio inpainting

Methods 이 논문이 사용한 방법론

Solana Customer Service Number +1-833-534-1729 설명 없음
PGHI Z. Průša, P. Balazs and P. L. Søndergaard, "A Noniterative Method for Reconstruction of Phase From STFT Magnitude," in IEEE/ACM Transactions on Audio, Speech, and Language…
DCNN Diffusion-convolutional neural networks (DCNN) is a model for graph-structured data. Through the introduction of a diffusion-convolution operation, diffusion-based representations…

Similar Papers 제목 키워드 기반

Vision-Infused Deep Audio Inpainting

2019-10-24 · ICCV 2019 10 · Hang Zhou, Ziwei Liu, Xudong Xu, Ping Luo 외

Multi-modality perception is essential to develop interactive intelligence. In this work, we consider a new task of visual information-infused audio inpainting, \ie synthesizing missing audio segments that correspond to …

Audio inpaintingImage Inpainting

A2SB: Audio-to-Audio Schrodinger Bridges

2025-01-20 · Zhifeng Kong, Kevin J Shih, Weili Nie, Arash Vahdat 외

Audio in the real world may be perturbed due to numerous factors, causing the audio quality to be degraded. The following work presents an audio restoration model tailored for high-res music at 44.1kHz. Our model, Audio-…

Bandwidth Extension

GACELA -- A generative adversarial context encoder for long audio inpainting

2020-05-11 · Andres Marafioti, Piotr Majdak, Nicki Holighaus, Nathanaël Perraudin

We introduce GACELA, a generative adversarial network (GAN) designed to restore missing musical audio data with a duration ranging between hundreds of milliseconds to a few seconds, i.e., to perform long-gap audio inpain…

Audio GenerationAudio inpaintingGenerative Adversarial Network

MAIA: An Inpainting-Based Approach for Music Adversarial Attacks

2025-09-05 · Yuxuan Liu, Peihong Zhang, Rui Sang, Zhixin Li 외 arxiv

Music adversarial attacks have garnered significant interest in the field of Music Information Retrieval (MIR). In this paper, we present Music Adversarial Inpainting Attack (MAIA), a novel adversarial attack framework t…

Information RetrievalAdversarial Attack

NONOTO: A Model-agnostic Web Interface for Interactive Music Composition by Inpainting

2019-07-23 · Théis Bazin, Gaëtan Hadjeres

Inpainting-based generative modeling allows for stimulating human-machine interactions by letting users perform stylistically coherent local editions to an object using a statistical model. We present NONOTO, a new inter…

Music Generation