paper-with-me

홈 › Papers

EnvGAN: Adversarial Synthesis of Environmental Sounds for Data Augmentation

2021-04-15 · Aswathy Madhu, Suresh K

The research in Environmental Sound Classification (ESC) has been progressively growing with the emergence of deep learning algorithms. However, data scarcity poses a major hurdle for any huge advance in this domain. Data augmentation offers an excellent solution to this problem. While Generative Adversarial Networks (GANs) have been successful in generating synthetic speech and sounds of musical instruments, they have hardly been applied to the generation of environmental sounds. This paper presents EnvGAN, the first ever application of GANs for the adversarial generation of environmental sounds. Our experiments on three standard ESC datasets illustrate that the EnvGAN can synthesize audio similar to the ones in the datasets. The suggested method of augmentation outshines most of the futuristic techniques for audio augmentation.

📄 PDF Abstract BibTeX arXiv:2104.07326

Code (0)

등록된 구현이 없습니다.

Tasks

Data AugmentationEnvironmental Sound ClassificationSound Classification

Similar Papers 제목 키워드 기반

A General Framework for Learning Procedural Audio Models of Environmental Sounds

2023-03-04 · Danzel Serrano, Mark Cartwright

This paper introduces the Procedural (audio) Variational autoEncoder (ProVE) framework as a general approach to learning Procedural Audio PA models of environmental sounds with an improvement to the realism of the synthe…

FAD

CAESynth: Real-Time Timbre Interpolation and Pitch Control with Conditional Autoencoders

2021-11-09 · IEEE MLSP 2021 9 · Aaron Valero Puche, Sukhan Lee

In this paper, we present a novel audio synthesizer, CAESynth, based on a conditional autoencoder. CAESynth synthesizes timbre in real-time by interpolating the reference sounds in their shared latent feature space, whil…

Audio SynthesisMixed RealityPitch controlTimbre Interpolation

DrumGAN VST: A Plugin for Drum Sound Analysis/Synthesis With Autoencoding Generative Adversarial Networks

2022-06-29 · Javier Nistal, Cyran Aouameur, Ithan Velarde, Stefan Lattner

In contemporary popular music production, drum sound design is commonly performed by cumbersome browsing and processing of pre-recorded samples in sound libraries. One can also use specialized synthesis hardware, typical…

Generative Adversarial NetworkResynthesis

Detection of Adversarial Attacks and Characterization of Adversarial Subspace

2019-10-26 · Mohammad Esmaeilpour, Patrick Cardinal, Alessandro Lameiras Koerich

Adversarial attacks have always been a serious threat for any data-driven model. In this paper, we explore subspaces of adversarial examples in unitary vector domain, and we propose a novel detector for defending our mod…

BenchmarkingEnvironmental Sound ClassificationregressionSound Classification

MTCRNN: A multi-scale RNN for directed audio texture synthesis

2020-11-25 · M. Huzaifah, L. Wyse

Audio textures are a subset of environmental sounds, often defined as having stable statistical characteristics within an adequately large window of time but may be unstructured locally. They include common everyday soun…

Texture Synthesis