paper-with-me

홈 › Papers

The Orchive : Data mining a massive bioacoustic archive

2013-07-02 · Steven Ness, Helena Symonds, Paul Spong, George Tzanetakis

The Orchive is a large collection of over 20,000 hours of audio recordings from the OrcaLab research facility located off the northern tip of Vancouver Island. It contains recorded orca vocalizations from the 1980 to the present time and is one of the largest resources of bioacoustic data in the world. We have developed a web-based interface that allows researchers to listen to these recordings, view waveform and spectral representations of the audio, label clips with annotations, and view the results of machine learning classifiers based on automatic audio features extraction. In this paper we describe such classifiers that discriminate between background noise, orca calls, and the voice notes that are present in most of the tapes. Furthermore we show classification results for individual calls based on a previously existing orca call catalog. We have also experimentally investigated the scalability of classifiers over the entire Orchive.

📄 PDF Abstract BibTeX arXiv:1307.0589

Code (1)

sness/orchive

Similar Papers 제목 키워드 기반

Transferable Models for Bioacoustics with Human Language Supervision

2023-08-09 · David Robinson, Adelaide Robinson, Lily Akrapongpisak

Passive acoustic monitoring offers a scalable, non-invasive method for tracking global biodiversity and anthropogenic impacts on species. Although deep learning has become a vital tool for processing this data, current m…

Masked Autoencoders with Limited Data: Does It Work? A Fine-Grained Bioacoustics Case Study

2026-05-13 · Wuao Liu, Mustafa Chasmai, Subhransu Maji, Grant Van Horn arxiv

Bioacoustic recognition requires fine-grained acoustic understanding to distinguish similar-sounding species. However, many large-scale data repositories such as iNaturalist are weakly annotated, often with only a single…

Self-Supervised Learning

Comparing Self-Supervised Learning Models Pre-Trained on Human Speech and Animal Vocalizations for Bioacoustics Processing

2025-01-10 · Eklavya Sarkar, Mathew Magimai. -Doss

Self-supervised learning (SSL) foundation models have emerged as powerful, domain-agnostic, general-purpose feature extractors applicable to a wide range of tasks. Such models pre-trained on human speech have demonstrate…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Self-Supervised Learningspeech-recognition+1

ArcGPT: A Large Language Model Tailored for Real-world Archival Applications

2023-07-27 · Shitou Zhang, Jingrui Hou, Siyuan Peng, Zuchao Li 외

Archives play a crucial role in preserving information and knowledge, and the exponential growth of such data necessitates efficient and automated tools for managing and utilizing archive information resources. Archival …

Language ModelingLanguage ModellingLarge Language ModelManagement

Foundation Models for Bioacoustics -- a Comparative Review

2025-08-02 · Raphael Schwinger, Paria Vali Zadeh, Lukas Rauch, Mats Kurz 외 arxiv

Automated bioacoustic analysis is essential for biodiversity monitoring and conservation, requiring advanced deep learning models that can adapt to diverse bioacoustic tasks. This article presents a comprehensive review …

Self-Supervised LearningRepresentation Learning