paper-with-me

Papers

Sub-Spectrogram Segmentation for Environmental Sound Classification via Convolutional Recurrent Neural Network and Score Level Fusion

2019-08-16 · Tianhao Qiao, Shunqing Zhang, Zhichao Zhang, Shan Cao, Shugong Xu

Environmental Sound Classification (ESC) is an important and challenging problem, and feature representation is a critical and even decisive factor in ESC. Feature representation ability directly affects the accuracy of sound classification. Therefore, the ESC performance is heavily dependent on the effectiveness of representative features extracted from the environmental sounds. In this paper, we propose a subspectrogram segmentation based ESC classification framework. In addition, we adopt the proposed Convolutional Recurrent Neural Network (CRNN) and score level fusion to jointly improve the classification accuracy. Extensive truncation schemes are evaluated to find the optimal number and the corresponding band ranges of sub-spectrograms. Based on the numerical experiments, the proposed framework can achieve 81.9% ESC classification accuracy on the public dataset ESC-50, which provides 9.1% accuracy improvement over traditional baseline schemes.

📄 PDF Abstract BibTeX arXiv:1908.05863

Code (0)

등록된 구현이 없습니다.

Tasks

ClassificationEnvironmental Sound ClassificationGeneral ClassificationSound Classification

Similar Papers 제목 키워드 기반

Unsupervised Feature Learning for Environmental Sound Classification Using Weighted Cycle-Consistent Generative Adversarial Network

2019-04-08 · Mohammad Esmaeilpour, Patrick Cardinal, Alessandro Lameiras Koerich

In this paper we propose a novel environmental sound classification approach incorporating unsupervised feature learning from codebook via spherical $K$-Means++ algorithm and a new architecture for high-level data augmen…

BenchmarkingClassificationData AugmentationEnvironmental Sound Classification+3

Masked Conditional Neural Networks for Environmental Sound Classification

2018-05-25 · Fady Medhat, David Chesmore, John Robinson

The ConditionaL Neural Network (CLNN) exploits the nature of the temporal sequencing of the sound signal represented in a spectrogram, and its variant the Masked ConditionaL Neural Network (MCLNN) induces the network to …

ClassificationEnvironmental Sound ClassificationGeneral ClassificationSound Classification

From Sound Representation to Model Robustness

2020-07-27 · Mohammad Esmaeilpour, Patrick Cardinal, Alessandro Lameiras Koerich

In this paper, we investigate the impact of different standard environmental sound representations (spectrograms) on the recognition performance and adversarial attack robustness of a victim residual convolutional neural…

Adversarial AttackAdversarial RobustnessBenchmarkingmodel

Utilizing Domain Knowledge in End-to-End Audio Processing

2017-12-01 · Tycho Max Sylvester Tax, Jose Luis Diez Antich, Hendrik Purwins, Lars Maaløe

End-to-end neural network based approaches to audio modelling are generally outperformed by models trained on high-level data representations. In this paper we present preliminary work that shows the feasibility of train…

Environmental Sound ClassificationGeneral ClassificationSound Classification

Spectral and Rhythm Feature Performance Evaluation for Category and Class Level Audio Classification with Deep Convolutional Neural Networks

2025-09-09 · Friedrich Wolf-Monheim arxiv

Next to decision tree and k-nearest neighbours algorithms deep convolutional neural networks (CNNs) are widely used to classify audio data in many domains like music, speech or environmental sounds. To train a specific C…

Audio Classification