Self-Supervised Learning for Few-Shot Bird Sound Classification
Self-supervised learning (SSL) in audio holds significant potential across various domains, particularly in situations where abundant, unlabeled data is readily available at no cost. This is pertinent in bioacoustics, where biologists routinely collect extensive sound datasets from the natural environment. In this study, we demonstrate that SSL is capable of acquiring meaningful representations of bird sounds from audio recordings without the need for annotations. Our experiments showcase that these learned representations exhibit the capacity to generalize to new bird species in few-shot learning (FSL) scenarios. Additionally, we show that selecting windows with high bird activation for self-supervised learning, using a pretrained audio neural network, significantly enhances the quality of the learned representations.
Code (1)
Tasks
ClassificationFew-Shot LearningSelf-Supervised LearningSound ClassificationSimilar Papers 제목 키워드 기반
Can Masked Autoencoders Also Listen to Birds?
Masked Autoencoders (MAEs) have shown competitive results in audio classification by learning rich semantic representations through an efficient self-supervised reconstruction task. However, general-purpose models fail t…
Audio ClassificationMulti-Label ClassificationMUlTI-LABEL-ClASSIFICATIONSelf-Supervised Learning+1Channel-Spatial-Based Few-Shot Bird Sound Event Detection
In this paper, we propose a model for bird sound event detection that focuses on a small number of training samples within the everyday long-tail distribution. As a result, we investigate bird sound detection using the f…
Event DetectionFew-Shot LearningSound ClassificationSound Event DetectionWeakly-Supervised Classification and Detection of Bird Sounds in the Wild.
It is easier to hear birds than see them, however, they still play an essential role in nature and they are excellent indicators of deteriorating environmental quality and pollution. Recent advances in Machine Learning a…
Audio ClassificationAudio TaggingBird Audio DetectionSound Event Detection+1Unsupervised classification to improve the quality of a bird song recording dataset
Open audio databases such as Xeno-Canto are widely used to build datasets to explore bird song repertoire or to train models for automatic bird sound classification by deep learning algorithms. However, such databases su…
Sound ClassificationTemporal LocalizationFew-shot Long-Tailed Bird Audio Recognition
It is easier to hear birds than see them. However, they still play an essential role in nature and are excellent indicators of deteriorating environmental quality and pollution. Recent advances in Deep Neural Networks al…