paper-with-me

홈 › Papers

A general-purpose deep learning approach to model time-varying audio effects

2019-05-15 · Marco A. Martínez Ramírez, Emmanouil Benetos, Joshua D. Reiss

Audio processors whose parameters are modified periodically over time are often referred as time-varying or modulation based audio effects. Most existing methods for modeling these type of effect units are often optimized to a very specific circuit and cannot be efficiently generalized to other time-varying effects. Based on convolutional and recurrent neural networks, we propose a deep learning architecture for generic black-box modeling of audio processors with long-term memory. We explore the capabilities of deep neural networks to learn such long temporal dependencies and we show the network modeling various linear and nonlinear, time-varying and time-invariant audio effects. In order to measure the performance of the model, we propose an objective metric based on the psychoacoustics of modulation frequency perception. We also analyze what the model is actually learning and how the given task is accomplished.

📄 PDF Abstract BibTeX arXiv:1905.06148

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Modeling of nonlinear audio effects with end-to-end deep neural networks

2018-10-15 · Marco Martínez, Joshua D. Reiss

In the context of music production, distortion effects are mainly used for aesthetic reasons and are usually applied to electric musical instruments. Most existing methods for nonlinear modeling are often either simplifi…

Modelling black-box audio effects with time-varying feature modulation

2022-11-01 · Marco Comunità, Christian J. Steinmetz, Huy Phan, Joshua D. Reiss

Deep learning approaches for black-box modelling of audio effects have shown promise, however, the majority of existing work focuses on nonlinear effects with behaviour on relatively short time-scales, such as guitar amp…

Time-Varying Audio Effect Modeling by End-to-End Adversarial Training

2025-12-17 · Yann Bourdin, Pierrick Legrand, Fanny Roche arxiv

Deep learning has become a standard approach for the modeling of audio effects, yet strictly black-box modeling remains problematic for time-varying systems. Unlike time-invariant effects, training models on devices with…

StepAudio 3 Gen Technical Report

2026-09-11 · Bin Lin, Bo Zhao, Boyang Wang, Boyang Zhang 외 hf

We introduce StepAudio 3 Gen, a general-purpose audio generation model that supports zero-shot text-to-speech (TTS), voice design, vocal generation, sound effects, music, vibe speech, and mixtures of multiple audio types…

Audio Generation

Modulation Extraction for LFO-driven Audio Effects

2023-05-22 · Christopher Mitcheltree, Christian J. Steinmetz, Marco Comunità, Joshua D. Reiss

Low frequency oscillator (LFO) driven audio effects such as phaser, flanger, and chorus, modify an input signal using time-varying filters and delays, resulting in characteristic sweeping or widening effects. It has been…