paper-with-me

Papers

AI-based soundscape analysis: Jointly identifying sound sources and predicting annoyance

2023-11-15 · Yuanbo Hou, Qiaoqiao Ren, Huizhong Zhang, Andrew Mitchell, Francesco Aletta, Jian Kang, Dick Botteldooren

Soundscape studies typically attempt to capture the perception and understanding of sonic environments by surveying users. However, for long-term monitoring or assessing interventions, sound-signal-based approaches are required. To this end, most previous research focused on psycho-acoustic quantities or automatic sound recognition. Few attempts were made to include appraisal (e.g., in circumplex frameworks). This paper proposes an artificial intelligence (AI)-based dual-branch convolutional neural network with cross-attention-based fusion (DCNN-CaF) to analyze automatic soundscape characterization, including sound recognition and appraisal. Using the DeLTA dataset containing human-annotated sound source labels and perceived annoyance, the DCNN-CaF is proposed to perform sound source classification (SSC) and human-perceived annoyance rating prediction (ARP). Experimental findings indicate that (1) the proposed DCNN-CaF using loudness and Mel features outperforms the DCNN-CaF using only one of them. (2) The proposed DCNN-CaF with cross-attention fusion outperforms other typical AI-based models and soundscape-related traditional machine learning methods on the SSC and ARP tasks. (3) Correlation analysis reveals that the relationship between sound sources and annoyance is similar for humans and the proposed AI-based DCNN-CaF model. (4) Generalization tests show that the proposed model's ARP in the presence of model-unknown sound sources is consistent with expert expectations and can explain previous findings from the literature on sound-scape augmentation.

📄 PDF Abstract BibTeX arXiv:2311.09030

Code (1)

yuanbo2020/ai-soundscape 공식 구현 pytorch

Similar Papers 제목 키워드 기반

Singapore Soundscape Site Selection Survey (S5): Identification of Characteristic Soundscapes of Singapore via Weighted k-means Clustering

2022-06-07 · Kenneth Ooi, Bhan Lam, Joo Young Hong, Karn N. Watcharasupat 외

The ecological validity of soundscape studies usually rests on a choice of soundscapes that are representative of the perceptual space under investigation. For example, a soundscape pleasantness study might investigate l…

Selection biasUnsupervised Spatial Clustering

Soundscape Captioning using Sound Affective Quality Network and Large Language Model

2024-06-09 · Yuanbo Hou, Qiaoqiao Ren, Andrew Mitchell, Wenwu Wang 외

We live in a rich and varied acoustic world, which is experienced by individuals or communities as a soundscape. Computational auditory scene analysis, disentangling acoustic scenes by detecting and classifying events, f…

Language ModelingLanguage ModellingLarge Language Model

Sat2Sound: A Unified Framework for Zero-Shot Soundscape Mapping

2025-05-19 · Subash Khanal, Srikumar Sastry, Aayush Dhakal, Adeel Ahmad 외

We present Sat2Sound, a multimodal representation learning framework for soundscape mapping, designed to predict the distribution of sounds at any location on Earth. Existing methods for this task rely on satellite image…

Contrastive LearningCross-Modal RetrievalDiversityImage Captioning+3

Scene2Sound: Auditory-Grounded Soundscape Generation for 3D Gaussian Worlds

2026-08-01 · Masaki Yoshida, Ren Togo, Takahiro Ogawa, Miki Haseyama arxiv

3D Gaussian Splatting (3DGS) turns captured or generated imagery into photorealistic 3D world simulations that users can freely explore, yet these worlds remain silent. Because existing audio generation methods condition…

Audio Generation

Deployment of an IoT System for Adaptive In-Situ Soundscape Augmentation

2022-04-29 · Trevor Wong, Karn N. Watcharasupat, Bhan Lam, Kenneth Ooi 외

Soundscape augmentation is an emerging approach for noise mitigation by introducing additional sounds known as "maskers" to increase acoustic comfort. Traditionally, the choice of maskers is often predicated on expert gu…

Cloud Computing