paper-with-me

Papers

Self-supervised Attention Model for Weakly Labeled Audio Event Classification

2019-08-07 · Bongjun Kim, Shabnam Ghaffarzadegan

We describe a novel weakly labeled Audio Event Classification approach based on a self-supervised attention model. The weakly labeled framework is used to eliminate the need for expensive data labeling procedure and self-supervised attention is deployed to help a model distinguish between relevant and irrelevant parts of a weakly labeled audio clip in a more effective manner compared to prior attention models. We also propose a highly effective strongly supervised attention model when strong labels are available. This model also serves as an upper bound for the self-supervised model. The performances of the model with self-supervised attention training are comparable to the strongly supervised one which is trained using strong labels. We show that our self-supervised attention method is especially beneficial for short audio events. We achieve 8.8% and 17.6% relative mean average precision improvements over the current state-of-the-art systems for SL-DCASE-17 and balanced AudioSet.

📄 PDF Abstract BibTeX arXiv:1908.02876

Code (0)

등록된 구현이 없습니다.

Tasks

ClassificationGeneral Classification

Similar Papers 제목 키워드 기반

Audio Event and Scene Recognition: A Unified Approach using Strongly and Weakly Labeled Data

2016-11-12 · Anurag Kumar, Bhiksha Raj

In this paper we propose a novel learning framework called Supervised and Weakly Supervised Learning where the goal is to learn simultaneously from weakly and strongly labeled data. Strongly labeled data can be simply un…

Scene RecognitionWeakly-supervised Learning

A Closer Look at Weak Label Learning for Audio Events

2018-04-24 · Ankit Shah, Anurag Kumar, Alexander G. Hauptmann, Bhiksha Raj

Audio content analysis in terms of sound events is an important research problem for a variety of applications. Recently, the development of weak labeling approaches for audio or sound event detection (AED) and availabil…

Audio ClassificationEvent DetectionSound Event DetectionWeakly-supervised Learning

Guided learning for weakly-labeled semi-supervised sound event detection

2019-06-06 · Liwei Lin, Xiangdong Wang, Hong Liu, Yueliang Qian

We propose a simple but efficient method termed Guided Learning for weakly-labeled semi-supervised sound event detection (SED). There are two sub-targets implied in weakly-labeled SED: audio tagging and boundary detectio…

Audio TaggingBoundary DetectionEvent DetectionGeneral Classification+1

Self-supervised Contrastive Learning for Audio-Visual Action Recognition

2022-04-28 · Yang Liu, Ying Tan, Haoyuan Lan

The underlying correlation between audio and visual modalities can be utilized to learn supervised information for unlabeled videos. In this paper, we propose an end-to-end self-supervised framework named Audio-Visual Co…

Action RecognitionContrastive LearningSelf-Supervised Action Recognition

CosyAudio: Improving Audio Generation with Confidence Scores and Synthetic Captions

2025-01-28 · Xinfa Zhu, Wenjie Tian, Xinsheng Wang, Lei He 외

Text-to-Audio (TTA) generation is an emerging area within AI-generated content (AIGC), where audio is created from natural language descriptions. Despite growing interest, developing robust TTA models remains challenging…

Audio captioningAudio Generation