Approximal operator with application to audio inpainting
In their recent evaluation of time-frequency representations and structured sparsity approaches to audio inpainting, Lieb and Stark (2018) have used a particular mapping as a proximal operator. This operator serves as the fundamental part of an iterative numerical solver. However, their mapping is improperly justified. The present article proves that their mapping is indeed a proximal operator, and also derives its proper counterpart. Furthermore, it is rationalized that Lieb and Stark's operator can be understood as an approximation of the proper mapping. Surprisingly, in most cases, such an approximation (referred to as the approximal operator) is shown to provide even better numerical results in audio inpainting compared to its proper counterpart, while being computationally much more effective.
Code (1)
Tasks
Audio inpaintingSimilar Papers 제목 키워드 기반
Vision-Infused Deep Audio Inpainting
Multi-modality perception is essential to develop interactive intelligence. In this work, we consider a new task of visual information-infused audio inpainting, \ie synthesizing missing audio segments that correspond to …
Audio inpaintingImage InpaintingDeep Video Inpainting Guided by Audio-Visual Self-Supervision
Humans can easily imagine a scene from auditory information based on their prior knowledge of audio-visual events. In this paper, we mimic this innate human ability in deep learning models to improve the quality of video…
audio-visual learningVideo InpaintingAudio-Visual Speech Inpainting with Deep Learning
In this paper, we present a deep-learning-based framework for audio-visual speech inpainting, i.e., the task of restoring the missing parts of an acoustic speech signal from reliable audio context and uncorrupted visual …
Deep LearningMulti-Task LearningVAInpaint: Zero-Shot Video-Audio inpainting framework with LLMs-driven Module
Video and audio inpainting for mixed audio-visual content has become a crucial task in multimedia editing recently. However, precisely removing an object and its corresponding audio from a video without affecting the res…
Video InpaintingGACELA -- A generative adversarial context encoder for long audio inpainting
We introduce GACELA, a generative adversarial network (GAN) designed to restore missing musical audio data with a duration ranging between hundreds of milliseconds to a few seconds, i.e., to perform long-gap audio inpain…
Audio GenerationAudio inpaintingGenerative Adversarial Network