Sound Event Detection
5개 벤치마크 · 논문 211편 · 이 태스크의 논문 보기 →
Benchmarks
Most implemented
Towards Deep Learning Models Resistant to Adversarial Attacks
WavCaps: A ChatGPT-Assisted Weakly-Labelled Audio Captioning Dataset for Audio-Language Multimodal Research
Effective Pre-Training of Audio Transformers for Sound Event Detection
Papers
Auto-AEG: Scalable Data Construction for Open-Vocabulary Audio Event Grounding
Large Audio-Language Models (LALMs) reason fluently about sound yet struggle to localize precisely when events occur, while classical Sound Event Detection attains frame-level precision only over a closed label set. At t…
Reinforcement LearningSound Event DetectionSemi-Supervised Sound Event Detection with Conditional Mixup and Embedding-Level Contrastive Loss
Sound event detection (SED) is a core module for acoustic environmental analysis, yet its performance is often limited by scarce labeled data. Recent systems leverage large pretrained audio foundation models, but effecti…
Sound Event DetectionContrastive LearningAn Analysis of Untrained Deep Reservoir Networks for Audio Surveillance
In this paper, we investigate untrained recurrent models from the Reservoir Computing (RC) paradigm for audio surveillance, focusing on bidirectional Echo State Networks with different depths, from shallow to deep config…
Computational EfficiencySound Event DetectionA Neuromorphic Trigger for Efficient Audio Event Detection
Efficient processing of continuous audio streams remains a key challenge for real-time and resource-constrained systems. This paper introduces a neuromorphic trigger for audio event detection, based on a spiking neural n…
Sound Event DetectionTowards Open World Sound Event Detection
Sound Event Detection (SED) plays a vital role in audio understanding, with applications in surveillance, smart cities, healthcare, and multimedia indexing. However, conventional SED systems operate under a closed-world …
Sound Event DetectionMMAudio-LABEL: Audio Event Labeling via Audio Generation for Silent Video
Recent advances in multimodal generation have enabled high-quality audio generation from silent videos. Practical applications, such as sound production, demand not only the generated audio but also explicit sound event …
Sound Event Detectionmultimodal generationAudio Generation