paper-with-me

홈 › Papers

Introduction to speech recognition

2024-02-01 · Gabriel Dauphin

This document contains lectures and practical experimentations using Matlab and implementing a system which is actually correctly classifying three words (one, two and three) with the help of a very small database. To achieve this performance, it uses speech modeling specificities, powerful computer algorithms (dynamic time warping and Dijktra's algorithm) and machine learning (nearest neighbor). This document introduces also some machine learning evaluation metrics.

📄 PDF Abstract BibTeX arXiv:2402.01778

Code (0)

등록된 구현이 없습니다.

Tasks

Dynamic Time Warpingspeech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

Cloud-based Automatic Speech Recognition Systems for Southeast Asian Languages

2022-10-07 · Lei Wang, Rong Tong, Cheung Chi Leung, Sunil Sivadas 외

This paper provides an overall introduction of our Automatic Speech Recognition (ASR) systems for Southeast Asian languages. As not much existing work has been carried out on such regional languages, a few difficulties s…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

BayesSpeech: A Bayesian Transformer Network for Automatic Speech Recognition

2023-01-16 · Will Rieger

Recent developments using End-to-End Deep Learning models have been shown to have near or better performance than state of the art Recurrent Neural Networks (RNNs) on Automatic Speech Recognition tasks. These models tend…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition+1

Bridging the Gap Between Monaural Speech Enhancement and Recognition with Distortion-Independent Acoustic Modeling

2019-03-11 · Peidong Wang, Ke Tan, DeLiang Wang

Monaural speech enhancement has made dramatic advances since the introduction of deep learning a few years ago. Although enhanced speech has been demonstrated to have better intelligibility and quality for human listener…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Speech Enhancementspeech-recognition+1

SememeASR: Boosting Performance of End-to-End Speech Recognition against Domain and Long-Tailed Data Shift with Sememe Semantic Knowledge

2023-09-04 · Jiaxu Zhu, Changhe Song, Zhiyong Wu, Helen Meng

Recently, excellent progress has been made in speech recognition. However, pure data-driven approaches have struggled to solve the problem in domain-mismatch and long-tailed data. Considering that knowledge-driven approa…

Domain Generalizationspeech-recognitionSpeech Recognition

Bangla-Wave: Improving Bangla Automatic Speech Recognition Utilizing N-gram Language Models

2022-09-13 · Mohammed Rakib, Md. Ismail Hossain, Nabeel Mohammed, Fuad Rahman

Although over 300M around the world speak Bangla, scant work has been done in improving Bangla voice-to-text transcription due to Bangla being a low-resource language. However, with the introduction of the Bengali Common…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Language ModelingLanguage Modelling+2