paper-with-me

홈 › Papers

A Generative Model for Natural Sounds Based on Latent Force Modelling

2018-02-02 · William J. Wilkinson, Joshua D. Reiss, Dan Stowell

Recent advances in analysis of subband amplitude envelopes of natural sounds have resulted in convincing synthesis, showing subband amplitudes to be a crucial component of perception. Probabilistic latent variable analysis is particularly revealing, but existing approaches don't incorporate prior knowledge about the physical behaviour of amplitude envelopes, such as exponential decay and feedback. We use latent force modelling, a probabilistic learning paradigm that incorporates physical knowledge into Gaussian process regression, to model correlation across spectral subband envelopes. We augment the standard latent force model approach by explicitly modelling correlations over multiple time steps. Incorporating this prior knowledge strengthens the interpretation of the latent functions as the source that generated the signal. We examine this interpretation via an experiment which shows that sounds generated by sampling from our probabilistic model are perceived to be more realistic than those generated by similar models based on nonnegative matrix factorisation, even in cases where our model is outperformed from a reconstruction error perspective.

📄 PDF Abstract BibTeX arXiv:1802.00680

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Sound Model Factory: An Integrated System Architecture for Generative Audio Modelling

2022-06-27 · Lonce Wyse, Purnima Kamath, Chitralekha Gupta

We introduce a new system for data-driven audio sound model design built around two different neural network architectures, a Generative Adversarial Network(GAN) and a Recurrent Neural Network (RNN), that takes advantage…

Generative Adversarial Network

Fast Timing-Conditioned Latent Audio Diffusion

2024-02-07 · Zach Evans, CJ Carr, Josiah Taylor, Scott H. Hawley 외

Generating long-form 44.1kHz stereo audio from text prompts can be computationally demanding. Further, most previous works do not tackle that music and sound effects naturally vary in their duration. Our research focuses…

Audio GenerationGPUText-to-Music Generation

An Integrated System Architecture for Generative Audio Modeling

2021-09-29 · Lonce Wyse, Purnima Kamath, Chitralekha Gupta

We introduce a new system for data-driven audio sound model design built around two different neural network architectures, a Generative Adversarial Network(GAN) and a Recurrent Neural Network (RNN), that takes advantage…

Generative Adversarial Network

Continuous descriptor-based control for deep audio synthesis

2023-02-27 · Ninon Devis, Nils Demerlé, Sarah Nabi, David Genova 외

Despite significant advances in deep models for music generation, the use of these techniques remains restricted to expert users. Before being democratized among musicians, generative models must first provide expressive…

Audio Synthesiscontinuous-controlContinuous ControlMusic Generation

Conditional Sound Generation Using Neural Discrete Time-Frequency Representation Learning

2021-07-21 · Xubo Liu, Turab Iqbal, Jinzheng Zhao, Qiushi Huang 외

Deep generative models have recently achieved impressive performance in speech and music synthesis. However, compared to the generation of those domain-specific sounds, generating general sounds (such as siren, gunshots)…

DiversityMusic GenerationRepresentation LearningSpeech Synthesis