paper-with-me

Papers

ODAS: Open embeddeD Audition System

2021-03-05 · François Grondin, Dominic Létourneau, Cédric Godin, Jean-Samuel Lauzon, Jonathan Vincent, Simon Michaud, Samuel Faucher, François Michaud

Artificial audition aims at providing hearing capabilities to machines, computers and robots. Existing frameworks in robot audition offer interesting sound source localization, tracking and separation performance, although involve a significant amount of computations that limit their use on robots with embedded computing capabilities. This paper presents ODAS, the Open embeddeD Audition System framework, which includes strategies to reduce the computational load and perform robot audition tasks on low-cost embedded computing systems. It presents key features of ODAS, along with cases illustrating its uses in different robots and artificial audition applications.

📄 PDF Abstract BibTeX arXiv:2103.03954

Code (1)

introlab/odas

Tasks

Sound Source Localization

Similar Papers 제목 키워드 기반

An embedded multichannel sound acquisition system for drone audition

2021-01-17 · Michael Clayton, Lin Wang, Andrew McPherson, Andrea Cavallaro

Microphone array techniques can improve the acoustic sensing performance on drones, compared to the use of a single microphone. However, multichannel sound acquisition systems are not available in current commercial dron…

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning

2025-05-31 · Yifan Peng, Shakeel Muhammad, Yui Sudo, William Chen 외

The Open Whisper-style Speech Models (OWSM) project has developed a series of fully open speech foundation models using academic-scale resources, but their training data remains insufficient. This work enhances OWSM by i…

Adaptive Robotic Arm Control with a Spiking Recurrent Neural Network on a Digital Accelerator

2024-05-21 · Alejandro Linares-Barranco, Luciano Prono, Robert Lengenstein, Giacomo Indiveri 외

With the rise of artificial intelligence, neural network simulations of biological neuron models are being explored to reduce the footprint of learning and inference in resource-constrained task scenarios. A mainstream t…

Automatic Detection and Annotation of Sperm Whale Codas

2024-07-24 · Guy Gubnitsky, Yaly Mevorach, Shane Gero, David F. Gruber 외

A key technology in sperm whale (Physeter macrocephalus) monitoring is the identification of sperm whale communication signals, known as codas. In this paper we present the first automatic coda detector and annotator. Th…

YODAS: Youtube-Oriented Dataset for Audio and Speech

2024-06-02 · Xinjian Li, Shinnosuke Takamichi, Takaaki Saeki, William Chen 외

In this study, we introduce YODAS (YouTube-Oriented Dataset for Audio and Speech), a large-scale, multilingual dataset comprising currently over 500k hours of speech data in more than 100 languages, sourced from both lab…

Self-Supervised Learningspeech-recognitionSpeech Recognition