paper-with-me

Papers

Acoustic Scene Classification using Audio Tagging

2020-04-19

Acoustic scene classification systems using deep neural networks classify given recordings into pre-defined classes. In this study, we propose a novel scheme for acoustic scene classification which adopts an audio tagging system inspired by the human perception mechanism. When humans identify an acoustic scene, the existence of different sound events provides discriminative information which affects the judgement. The proposed framework mimics this mechanism using various approaches. Firstly, we employ three methods to concatenate tag vectors extracted using an audio tagging system with an intermediate hidden layer of an acoustic scene classification system. We also explore the multi-head attention on the feature map of an acoustic scene classification system using tag vectors. Experiments conducted on the detection and classification of acoustic scenes and events 2019 task 1-a dataset demonstrate the effectiveness of the proposed scheme. Concatenation and multi-head attention show a classification accuracy of 75.66 % and 75.58 %, respectively, compared to 73.63 % accuracy of the baseline. The system with the proposed two approaches combined demonstrates an accuracy of 76.75 %.

📄 PDF Abstract BibTeX arXiv:2003.09164

Code (0)

등록된 구현이 없습니다.

Tasks

Acoustic Scene ClassificationAudio TaggingClassificationScene ClassificationTAG

Similar Papers 제목 키워드 기반

An evaluation of data augmentation methods for sound scene geotagging

2021-10-09 · Helen L. Bear, Veronica Morfi, Emmanouil Benetos

Sound scene geotagging is a new topic of research which has evolved from acoustic scene classification. It is motivated by the idea of audio surveillance. Not content with only describing a scene in a recording, a machin…

Acoustic Scene ClassificationClassificationData AugmentationScene Classification

Classifying Variable-Length Audio Files with All-Convolutional Networks and Masked Global Pooling

2016-07-11 · Lars Hertel, Huy Phan, Alfred Mertins

We trained a deep all-convolutional neural network with masked global pooling to perform single-label classification for acoustic scene classification and multi-label classification for domestic audio tagging in the DCAS…

Acoustic Scene ClassificationAllAudio TaggingClassification+5

Unsupervised Feature Learning Based on Deep Models for Environmental Audio Tagging

2016-07-13 · Yong Xu, Qiang Huang, Wenwu Wang, Peter Foster 외

Environmental audio tagging aims to predict only the presence or absence of certain acoustic events in the interested acoustic scene. In this paper we make contributions to audio tagging in two parts, respectively, acous…

Audio TaggingGeneral ClassificationMulti-Label ClassificationMUlTI-LABEL-ClASSIFICATION

Joint framework with deep feature distillation and adaptive focal loss for weakly supervised audio tagging and acoustic event detection

2021-03-23 · Yunhao Liang, Yanhua Long, Yijie Li, Jiaen Liang 외

A good joint training framework is very helpful to improve the performances of weakly supervised audio tagging (AT) and acoustic event detection (AED) simultaneously. In this study, we propose three methods to improve th…

Audio TaggingEvent Detection

SpectNet : End-to-End Audio Signal Classification Using Learnable Spectrograms

2022-11-17 · Md. Istiaq Ansari, Taufiq Hasan

Pattern recognition from audio signals is an active research topic encompassing audio tagging, acoustic scene classification, music classification, and other areas. Spectrogram and mel-frequency cepstral coefficients (MF…

Acoustic Scene ClassificationAnomaly DetectionAudio ClassificationAudio Tagging+4