paper-with-me

Papers

Do sound event representations generalize to other audio tasks? A case study in audio transfer learning

2021-06-21 · Anurag Kumar, Yun Wang, Vamsi Krishna Ithapu, Christian Fuegen

Transfer learning is critical for efficient information transfer across multiple related learning problems. A simple, yet effective transfer learning approach utilizes deep neural networks trained on a large-scale task for feature extraction. Such representations are then used to learn related downstream tasks. In this paper, we investigate transfer learning capacity of audio representations obtained from neural networks trained on a large-scale sound event detection dataset. We build and evaluate these representations across a wide range of other audio tasks, via a simple linear classifier transfer mechanism. We show that such simple linear transfer is already powerful enough to achieve high performance on the downstream tasks. We also provide insights into the attributes of sound event representations that enable such efficient information transfer.

📄 PDF Abstract BibTeX arXiv:2106.11335

Code (0)

등록된 구현이 없습니다.

Tasks

Event DetectionSound Event DetectionTransfer Learning

Similar Papers 제목 키워드 기반

Learning Multi-Target TDOA Features for Sound Event Localization and Detection

2024-08-30 · Axel Berg, Johanna Engman, Jens Gulin, Karl Åström 외

Sound event localization and detection (SELD) systems using audio recordings from a microphone array rely on spatial cues for determining the location of sound events. As a consequence, the localization performance of su…

Sound Event Localization and Detection

Hypernetworks build Implicit Neural Representations of Sounds

2023-02-09 · Filip Szatkowski, Karol J. Piczak, Przemysław Spurek, Jacek Tabor 외

Implicit Neural Representations (INRs) are nowadays used to represent multimedia signals across various real-life applications, including image super-resolution, image compression, or 3D rendering. Existing methods that …

Image CompressionImage Super-ResolutionMeta-LearningSuper-Resolution

What Do Language Models Hear? Probing for Auditory Representations in Language Models

2024-02-26 · Jerry Ngo, Yoon Kim

This work explores whether language models encode meaningfully grounded representations of sounds of objects. We learn a linear probe that retrieves the correct text representation of an object given a snippet of audio r…

Object

Learning Sound Events From Webly Labeled Data

2018-11-25 · 28th International Joint Conference on Artificial Intelligence 2019 8 · Anurag Kumar, Ankit Shah, Alex Hauptmann, Bhiksha Raj

In the last couple of years, weakly labeled learning for sound events has turned out to be an exciting approach for audio event detection. In this work, we introduce webly labeled learning for sound events in which we ai…

Event DetectionSound Event DetectionTransfer Learning

Binaural Signal Representations for Joint Sound Event Detection and Acoustic Scene Classification

2022-09-13 · Daniel Aleksander Krause, Annamaria Mesaros

Sound event detection (SED) and Acoustic scene classification (ASC) are two widely researched audio tasks that constitute an important part of research on acoustic scene analysis. Considering shared information between s…

Acoustic Scene ClassificationEvent DetectionScene ClassificationSound Event Detection