paper-with-me

Environmental Sound Classification

3개 벤치마크 · 논문 50편 · 이 태스크의 논문 보기 →

Benchmarks

UrbanSound8K

결과 9개

ESC-50

결과 4개

FSD50K

결과 3개

Most implemented

Papers

MAEB: Massive Audio Embedding Benchmark

2026-02-17 · Adnan El Assadi, Isaac Chung, Chenghao Xiao, Roman Solomatin 외 arxiv

We introduce the Massive Audio Embedding Benchmark (MAEB), a large-scale benchmark covering 30 tasks across speech, music, environmental sounds, and cross-modal audio-text reasoning in 100+ languages. We evaluate 50+ mod…

Environmental Sound Classification

Expressive Range Characterization of Open Text-to-Audio Models

2025-10-31 · Jonathan Morse, Azadeh Naderi, Swen Gaudl, Mark Cartwright 외 arxiv

Text-to-audio models are a type of generative model that produces audio output in response to a given textual prompt. Although level generators and the properties of the functional content that they create (e.g., playabi…

Environmental Sound Classification

Compressing Quaternion Convolutional Neural Networks for Audio Classification

2025-10-24 · Arshdeep Singh, Vinayak Abrol, Mark D. Plumbley arxiv

Conventional Convolutional Neural Networks (CNNs) in the real domain have been widely used for audio classification. However, their convolution operations process multi-channel inputs independently, limiting the ability …

Environmental Sound ClassificationSpeech Emotion RecognitionMusic Genre RecognitionKnowledge Distillation

ASDA: Audio Spectrogram Differential Attention Mechanism for Self-Supervised Representation Learning

2025-07-03 · Junyu Wang, Tianrui Wang, Meng Ge, Longbiao Wang 외 arxiv

In recent advancements in audio self-supervised representation learning, the standard Transformer architecture has emerged as the predominant approach, yet its attention mechanism often allocates a portion of attention w…

Environmental Sound ClassificationRepresentation LearningAudio ClassificationKeyword Spotting

Domain Adaptation Method and Modality Gap Impact in Audio-Text Models for Prototypical Sound Classification

2025-06-04 · Emiliano Acevedo, Martín Rocamora, Magdalena Fuentes

Audio-text models are widely used in zero-shot environmental sound classification as they alleviate the need for annotated data. However, we show that their performance severely drops in the presence of background sound …

ClassificationDomain AdaptationEnvironmental Sound ClassificationSound Classification

Weakly Supervised Convolutional Dictionary Learning for Multi-Label Classification

2025-03-11 · Hao Chen, Dayuan Tan

Convolutional Dictionary Learning (CDL) has emerged as a powerful approach for signal representation by learning translation-invariant features through convolution operations. While existing CDL methods are predominantly…

ClassificationDictionary LearningEnvironmental Sound ClassificationMulti-Label Classification+2

전체 50편 보기 →