paper-with-me

홈 › Papers

Multi-level graph learning for audio event classification and human-perceived annoyance rating prediction

2023-12-15 · Yuanbo Hou, Qiaoqiao Ren, Siyang Song, Yuxin Song, Wenwu Wang, Dick Botteldooren

WHO's report on environmental noise estimates that 22 M people suffer from chronic annoyance related to noise caused by audio events (AEs) from various sources. Annoyance may lead to health issues and adverse effects on metabolic and cognitive systems. In cities, monitoring noise levels does not provide insights into noticeable AEs, let alone their relations to annoyance. To create annoyance-related monitoring, this paper proposes a graph-based model to identify AEs in a soundscape, and explore relations between diverse AEs and human-perceived annoyance rating (AR). Specifically, this paper proposes a lightweight multi-level graph learning (MLGL) based on local and global semantic graphs to simultaneously perform audio event classification (AEC) and human annoyance rating prediction (ARP). Experiments show that: 1) MLGL with 4.1 M parameters improves AEC and ARP results by using semantic node information in local and global context aware graphs; 2) MLGL captures relations between coarse and fine-grained AEs and AR well; 3) Statistical analysis of MLGL results shows that some AEs from different sources significantly correlate with AR, which is consistent with previous research on human perception of these sound sources.

📄 PDF Abstract BibTeX arXiv:2312.09952

Code (1)

yuanbo2020/mlgl pytorch

Tasks

Graph Learning

Methods 이 논문이 사용한 방법론

AWARE We propose to theoretically and empirically examine the effect of incorporating weighting schemes into walk-aggregating GNNs. To this end, we propose a simple, interpretable, and…

Similar Papers 제목 키워드 기반

An Ontology-Aware Framework for Audio Event Classification

2020-01-27 · Yiwei Sun, Shabnam Ghaffarzadegan

Recent advancements in audio event classification often ignore the structure and relation between the label classes available as prior information. This structure can be defined by ontology and augmented in the classifie…

ClassificationGeneral Classification

Multi-level Attention Model for Weakly Supervised Audio Classification

2018-03-06 · Changsong Yu, Karim Said Barsim, Qiuqiang Kong, Bin Yang

In this paper, we propose a multi-level attention model to solve the weakly labelled audio classification problem. The objective of audio classification is to predict the presence or absence of audio events in an audio c…

Audio Classification

Multi-level Attention Fusion Network for Audio-visual Event Recognition

2021-06-12 · Mathilde Brousmiche, Jean Rouat, Stéphane Dupont

Event classification is inherently sequential and multimodal. Therefore, deep neural models need to dynamically focus on the most relevant time window and/or modality of a video. In this study, we propose the Multi-level…

Multi-dimensional Edge-based Audio Event Relational Graph Representation Learning for Acoustic Scene Classification

2022-10-27 · Yuanbo Hou, Siyang Song, Chuang Yu, Yuxin Song 외

Most existing deep learning-based acoustic scene classification (ASC) approaches directly utilize representations extracted from spectrograms to identify target scenes. However, these approaches pay little attention to t…

Acoustic Scene ClassificationGraph Representation LearningRepresentation LearningScene Classification

Audio Event-Relational Graph Representation Learning for Acoustic Scene Classification

2023-10-05 · Yuanbo Hou, Siyang Song, Chuang Yu, Wenwu Wang 외

Most deep learning-based acoustic scene classification (ASC) approaches identify scenes based on acoustic features converted from audio clips containing mixed information entangled by polyphonic audio events (AEs). Howev…

Acoustic Scene ClassificationGraph Representation LearningRepresentation LearningScene Classification