paper-with-me

Papers

City classification from multiple real-world sound scenes

2019-07-29

The majority of sound scene analysis work focuses on one of two clearly defined tasks: acoustic scene classification or sound event detection. Whilst this separation of tasks is useful for problem definition, they inherently ignore some subtleties of the real-world, in particular how humans vary in how they describe a scene. Some will describe the weather and features within it, others will use a holistic descriptor like `park', and others still will use unique identifiers such as cities or names. In this paper, we undertake the task of automatic city classification to ask whether we can recognize a city from a set of sound scenes? In this problem each city has recordings from multiple scenes. We test a series of methods for this novel task and show that a simple convolutional neural network (CNN) can achieve accuracy of 50%. This is less than the acoustic scene classification task baseline in the DCASE 2018 ASC challenge on the same data. A simple adaptation to the class labels of pairing city labels with grouped scenes, accuracy increases to 52%, closer to the simpler scene classification task. Finally we also formulate the problem in a multi-task learning framework and achieve an accuracy of 56%, outperforming the aforementioned approaches.

📄 PDF Abstract BibTeX arXiv:1905.00979

Code (1)

drylbear/soundscapeCityClassification 공식 구현

Tasks

Acoustic Scene ClassificationClassificationEvent DetectionMulti-Task LearningScene ClassificationSound Event Detection

Similar Papers 제목 키워드 기반

Generative Diffusion Model Bootstraps Zero-shot Classification of Fetal Ultrasound Images In Underrepresented African Populations

2024-07-29 · Fangyijie Wang, Kevin Whelan, Guénolé Silvestre, Kathleen M. Curran

Developing robust deep learning models for fetal ultrasound image analysis requires comprehensive, high-quality datasets to effectively learn informative data representations within the domain. However, the scarcity of l…

zero-shot-classificationZero-Shot Learning

Enemy Spotted: in-game gun sound dataset for gunshot classification and localization

2022-08-22 · IEEE Conference on Games (GoG) 2022 8 · Junwoo Park; Youngwoo Cho; Gyuhyeon Sim; Hojoon Lee; Jaegul Choo

Recently, deep learning-based methods have drawn huge attention due to their simple yet high performance without domain knowledge in sound classification and localization tasks. However, a lack of gun sounds in existing …

ClassificationSound Classification

Enemy Spotted: in-game gun sound dataset for gunshot classification and localization

2022-10-12 · Junwoo Park, Youngwoo Cho, Gyuhyeon Sim, Hojoon Lee 외

Recently, deep learning-based methods have drawn huge attention due to their simple yet high performance without domain knowledge in sound classification and localization tasks. However, a lack of gun sounds in existing …

ClassificationSound Classification

SONYC-UST-V2: An Urban Sound Tagging Dataset with Spatiotemporal Context

2020-09-11 · Mark Cartwright, Jason Cramer, Ana Elisa Mendez Mendez, Yu Wang 외

We present SONYC-UST-V2, a dataset for urban sound tagging with spatiotemporal information. This dataset is aimed for the development and evaluation of machine listening systems for real-world urban noise monitoring. Whi…

Heterogeneous sound classification with the Broad Sound Taxonomy and Dataset

2024-10-01 · Panagiota Anastasopoulou, Jessica Torrey, Xavier Serra, Frederic Font

Automatic sound classification has a wide range of applications in machine listening, enabling context-aware sound processing and understanding. This paper explores methodologies for automatically classifying heterogeneo…

ClassificationSound Classification