paper-with-me

홈 › Papers

Masked Conditional Neural Networks for Environmental Sound Classification

2018-05-25 · Fady Medhat, David Chesmore, John Robinson

The ConditionaL Neural Network (CLNN) exploits the nature of the temporal sequencing of the sound signal represented in a spectrogram, and its variant the Masked ConditionaL Neural Network (MCLNN) induces the network to learn in frequency bands by embedding a filterbank-like sparseness over the network's links using a binary mask. Additionally, the masking automates the exploration of different feature combinations concurrently analogous to handcrafting the optimum combination of features for a recognition task. We have evaluated the MCLNN performance using the Urbansound8k dataset of environmental sounds. Additionally, we present a collection of manually recorded sounds for rail and road traffic, YorNoise, to investigate the confusion rates among machine generated sounds possessing low-frequency components. MCLNN has achieved competitive results without augmentation and using 12% of the trainable parameters utilized by an equivalent model based on state-of-the-art Convolutional Neural Networks on the Urbansound8k. We extended the Urbansound8k dataset with YorNoise, where experiments have shown that common tonal properties affect the classification performance.

📄 PDF Abstract BibTeX arXiv:1805.10004

Code (2)

fadymedhat/MCLNN 공식 구현 tf
fadymedhat/YorNoise 공식 구현

Tasks

ClassificationEnvironmental Sound ClassificationGeneral ClassificationSound Classification

Similar Papers 제목 키워드 기반

Environmental Sound Recognition using Masked Conditional Neural Networks

2018-04-08 · Fady Medhat, David Chesmore, John Robinson

Neural network based architectures used for sound recognition are usually adapted from other application domains, which may not harness sound related properties. The ConditionaL Neural Network (CLNN) is designed to consi…

Masked Conditional Neural Networks for Automatic Sound Events Recognition

2018-02-15 · Fady Medhat, David Chesmore, John Robinson

Deep neural network architectures designed for application domains other than sound, especially image recognition, may not optimally harness the time-frequency representation when adapted to the sound recognition problem…

Recognition of Acoustic Events Using Masked Conditional Neural Networks

2018-02-07 · Fady Medhat, David Chesmore, John Robinson

Automatic feature extraction using neural networks has accomplished remarkable success for images, but for sound recognition, these models are usually modified to fit the nature of the multi-dimensional temporal represen…

Deep Convolutional Neural Networks and Data Augmentation for Environmental Sound Classification

2017-01-23 · IEEE Signal Processing Letters 2017 1 · J. Salamon, J. P. Bello

The ability of deep convolutional neural networks (CNNs) to learn discriminative spectro-temporal patterns makes them well suited to environmental sound classification. However, the relative scarcity of labeled data has …

ClassificationData AugmentationDictionary LearningEnvironmental Sound Classification+1

Deep Convolutional Neural Networks and Data Augmentation for Environmental Sound Classification

2016-08-15 · IEEE Signal Processing Letters 2017 1 · Justin Salamon, Juan Pablo Bello

The ability of deep convolutional neural networks (CNN) to learn discriminative spectro-temporal patterns makes them well suited to environmental sound classification. However, the relative scarcity of labeled data has i…

Data AugmentationDictionary LearningEnvironmental Sound ClassificationGeneral Classification+1