paper-with-me

Papers

Semi-Supervised Audio Representation Learning for Modeling Beehive Strengths

2021-05-21 · Tony Zhang, Szymon Zmyslony, Sergei Nozdrenkov, Matthew Smith, Brandon Hopkins

Honey bees are critical to our ecosystem and food security as a pollinator, contributing 35% of our global agriculture yield. In spite of their importance, beekeeping is exclusively dependent on human labor and experience-derived heuristics, while requiring frequent human checkups to ensure the colony is healthy, which can disrupt the colony. Increasingly, pollinator populations are declining due to threats from climate change, pests, environmental toxicity, making their management even more critical than ever before in order to ensure sustained global food security. To start addressing this pressing challenge, we developed an integrated hardware sensing system for beehive monitoring through audio and environment measurements, and a hierarchical semi-supervised deep learning model, composed of an audio modeling module and a predictor, to model the strength of beehives. The model is trained jointly on audio reconstruction and prediction losses based on human inspections, in order to model both low-level audio features and circadian temporal dynamics. We show that this model performs well despite limited labels, and can learn an audio embedding that is useful for characterizing different sound profiles of beehives. This is the first instance to our knowledge of applying audio-based deep learning to model beehives and population size in an observational setting across a large number of hives.

📄 PDF Abstract BibTeX arXiv:2105.10536

Code (0)

등록된 구현이 없습니다.

Tasks

ManagementRepresentation Learning

Similar Papers 제목 키워드 기반

UrBAN: Urban Beehive Acoustics and PheNotyping Dataset

2024-06-05 · Mahsa Abdollahi, Yi Zhu, Heitor R. Guimarães, Nico Coallier 외

In this paper, we present a multimodal dataset obtained from a honey bee colony in Montr\'eal, Quebec, Canada, spanning the years of 2021 to 2022. This apiary comprised 10 beehives, with microphones recording more than 2…

Semi-Supervised Diseased Detection from Speech Dialogues with Multi-Level Data Modeling

2026-01-08 · Xingyuan Li, Mengyue Wu arxiv

Detecting medical conditions from speech acoustics is fundamentally a weakly-supervised learning problem: a single, often noisy, session-level label must be linked to nuanced patterns within a long, complex audio recordi…

End-to-end ASR: from Supervised to Semi-Supervised Learning with Modern Architectures

2019-11-19 · Gabriel Synnaeve, Qiantong Xu, Jacob Kahn, Tatiana Likhomanenko 외

We study pseudo-labeling for the semi-supervised training of ResNet, Time-Depth Separable ConvNets, and Transformers for speech recognition, with either CTC or Seq2Seq loss functions. We perform experiments on the standa…

Language ModelingLanguage Modellingspeech-recognitionSpeech Recognition

Anomalous Sound Detection using unsupervised and semi-supervised autoencoders and gammatone audio representation

2020-06-27 · Sergi Perez-Castanos, Javier Naranjo-Alcazar, Pedro Zuccarello, Maximo Cobos

Anomalous sound detection (ASD) is, nowadays, one of the topical subjects in machine listening discipline. Unsupervised detection is attracting a lot of interest due to its immediate applicability in many fields. For exa…

QS-TTS: Towards Semi-Supervised Text-to-Speech Synthesis via Vector-Quantized Self-Supervised Speech Representation Learning

2023-08-31 · Haohan Guo, Fenglong Xie, Jiawen Kang, Yujia Xiao 외

This paper proposes a novel semi-supervised TTS framework, QS-TTS, to improve TTS quality with lower supervised data requirements via Vector-Quantized Self-Supervised Speech Representation Learning (VQ-S3RL) utilizing mo…

Representation LearningSpeech Representation LearningSpeech Synthesistext-to-speech+3