paper-with-me

홈 › Papers

Deep Spiking Neural Networks for Large Vocabulary Automatic Speech Recognition

2019-11-19 · Jibin Wu, Emre Yilmaz, Malu Zhang, Haizhou Li, Kay Chen Tan

Artificial neural networks (ANN) have become the mainstream acoustic modeling technique for large vocabulary automatic speech recognition (ASR). A conventional ANN features a multi-layer architecture that requires massive amounts of computation. The brain-inspired spiking neural networks (SNN) closely mimic the biological neural networks and can operate on low-power neuromorphic hardware with spike-based computation. Motivated by their unprecedented energyefficiency and rapid information processing capability, we explore the use of SNNs for speech recognition. In this work, we use SNNs for acoustic modeling and evaluate their performance on several large vocabulary recognition scenarios. The experimental results demonstrate competitive ASR accuracies to their ANN counterparts, while require significantly reduced computational cost and inference time. Integrating the algorithmic power of deep SNNs with energy-efficient neuromorphic hardware, therefore, offer an attractive solution for ASR applications running locally on mobile and embedded devices.

📄 PDF Abstract BibTeX arXiv:1911.08373

Code (1)

deepspike/snn-for-asr pytorch

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

Surrogate Gradient Spiking Neural Networks as Encoders for Large Vocabulary Continuous Speech Recognition

2022-12-01 · Alexandre Bittar, Philip N. Garner

Compared to conventional artificial neurons that produce dense and real-valued responses, biologically-inspired spiking neurons transmit sparse and binary information, which can also lead to energy-efficient implementati…

speech-recognitionSpeech Recognition

Tigrinya Automatic Speech recognition with Morpheme based recognition units

2020-07-01 · WS 2020 7 · Hafte Abera, sebsibe hailemariam

The Tigrinya language is agglutinative and has a large number of inflected and derived forms of words. Therefore a Tigrinya large vocabulary continuous speech recognition system often has a large number of different unit…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Language ModelingLanguage Modelling+3

Automatic Speech Recognition with Very Large Conversational Finnish and Estonian Vocabularies

2017-07-13 · Seppo Enarvi, Peter Smit, Sami Virpioja, Mikko Kurimo

Today, the vocabulary size for language models in large vocabulary speech recognition is typically several hundreds of thousands of words. While this is already sufficient in some applications, the out-of-vocabulary word…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Large Vocabulary Spontaneous Speech Recognition for Tigrigna

2023-10-15 · Ataklti Kahsu, Solomon Teferra

This thesis proposes and describes a research attempt at designing and developing a speaker independent spontaneous automatic speech recognition system for Tigrigna The acoustic model of the Speech Recognition System is …

Automatic Speech RecognitionLanguage ModelingLanguage Modellingspeech-recognition+1

Complex Dynamic Neurons Improved Spiking Transformer Network for Efficient Automatic Speech Recognition

2023-02-02 · Minglun Han, Qingyu Wang, Tielin Zhang, Yi Wang 외

The spiking neural network (SNN) using leaky-integrated-and-fire (LIF) neurons has been commonly used in automatic speech recognition (ASR) tasks. However, the LIF neuron is still relatively simple compared to that in th…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition