paper-with-me

Papers

General-purpose Tagging of Freesound Audio with AudioSet Labels: Task Description, Dataset, and Baseline

2018-07-26 · Eduardo Fonseca, Manoj Plakal, Frederic Font, Daniel P. W. Ellis, Xavier Favory, Jordi Pons, Xavier Serra

This paper describes Task 2 of the DCASE 2018 Challenge, titled "General-purpose audio tagging of Freesound content with AudioSet labels". This task was hosted on the Kaggle platform as "Freesound General-Purpose Audio Tagging Challenge". The goal of the task is to build an audio tagging system that can recognize the category of an audio clip from a subset of 41 diverse categories drawn from the AudioSet Ontology. We present the task, the dataset prepared for the competition, and a baseline system.

📄 PDF Abstract BibTeX arXiv:1807.09902

Code (3)

asudomoeva/Audio-Tagging tf
dbalaji5/audiotagging tf
iooops/CS221-Audio-Tagging

Tasks

Audio TaggingTask 2

Similar Papers 제목 키워드 기반

FSD50K: An Open Dataset of Human-Labeled Sound Events

2020-10-01 · Eduardo Fonseca, Xavier Favory, Jordi Pons, Frederic Font 외

Most existing datasets for sound event recognition (SER) are relatively small and/or domain-specific, with the exception of AudioSet, based on over 2M tracks from YouTube videos and encompassing over 500 sound classes. H…

Combining High-Level Features of Raw Audio Waves and Mel-Spectrograms for Audio Tagging

2018-11-26 · Marcel Lederle, Benjamin Wilhelm

In this paper, we describe our contribution to Task 2 of the DCASE 2018 Audio Challenge. While it has become ubiquitous to utilize an ensemble of machine learning methods for classification tasks to obtain better predict…

Audio TaggingData AugmentationGeneral ClassificationTask 2

PSLA: Improving Audio Tagging with Pretraining, Sampling, Labeling, and Aggregation

2021-02-02 · Yuan Gong, Yu-An Chung, James Glass

Audio tagging is an active research area and has a wide range of applications. Since the release of AudioSet, great progress has been made in advancing model performance, which mostly comes from the development of novel …

Audio ClassificationAudio TaggingData AugmentationGeneral Classification

Improved Zero-Shot Audio Tagging & Classification with Patchout Spectrogram Transformers

2022-08-24 · Paul Primus, Gerhard Widmer

Standard machine learning models for tagging and classifying acoustic signals cannot handle classes that were not seen during training. Zero-Shot (ZS) learning overcomes this restriction by predicting classes based on ad…

Audio TaggingClassificationEnvironmental Sound ClassificationSound Classification

Audio tagging with noisy labels and minimal supervision

2019-06-07 · Eduardo Fonseca, Manoj Plakal, Frederic Font, Daniel P. W. Ellis 외

This paper introduces Task 2 of the DCASE2019 Challenge, titled "Audio tagging with noisy labels and minimal supervision". This task was hosted on the Kaggle platform as "Freesound Audio Tagging 2019". The task evaluates…

Audio TaggingTask 2