paper-with-me

Papers

Exploiting Hidden Representations from a DNN-based Speech Recogniser for Speech Intelligibility Prediction in Hearing-impaired Listeners

2022-04-08 · Zehai Tu, Ning Ma, Jon Barker

An accurate objective speech intelligibility prediction algorithms is of great interest for many applications such as speech enhancement for hearing aids. Most algorithms measures the signal-to-noise ratios or correlations between the acoustic features of clean reference signals and degraded signals. However, these hand-picked acoustic features are usually not explicitly correlated with recognition. Meanwhile, deep neural network (DNN) based automatic speech recogniser (ASR) is approaching human performance in some speech recognition tasks. This work leverages the hidden representations from DNN-based ASR as features for speech intelligibility prediction in hearing-impaired listeners. The experiments based on a hearing aid intelligibility database show that the proposed method could make better prediction than a widely used short-time objective intelligibility (STOI) based binaural measure.

📄 PDF Abstract BibTeX arXiv:2204.04287

Code (1)

claritychallenge/clarity/tree/main/recipes/cpc1/e032_sheffield 공식 구현

Tasks

PredictionSpeech Enhancementspeech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

The EASR Corpora of European Portuguese, French, Hungarian and Polish Elderly Speech

2014-05-01 · LREC 2014 5 · Annika H{\"a}m{\"a}l{\"a}inen, Jairo Avelar, Silvia Rodrigues, Miguel Sales Dias 외

Currently available speech recognisers do not usually work well with elderly speech. This is because several characteristics of speech (e.g. fundamental frequency, jitter, shimmer and harmonic noise ratio) change with ag…

Speech Recognition

Free on-line speech recogniser based on Kaldi ASR toolkit producing word posterior lattices

2014-06-01 · WS 2014 6 · Ond{\v{r}}ej Pl{\'a}tek, Filip Jur{\v{c}}{\'\i}{\v{c}}ek
Acoustic ModellingLanguage ModellingSpeech Recognition

Confidence Estimation for Attention-based Sequence-to-sequence Models for Speech Recognition

2020-10-22 · Qiujia Li, David Qiu, Yu Zhang, Bo Li 외

For various speech-related tasks, confidence scores from a speech recogniser are a useful measure to assess the quality of transcriptions. In traditional hidden Markov model-based automatic speech recognition (ASR) syste…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)DecoderLanguage Modeling+3

A Shared Task for Spoken CALL?

2016-05-01 · LREC 2016 5 · Claudia Baur, Johanna Gerlach, Manny Rayner, Martin Russell 외

We argue that the field of spoken CALL needs a shared task in order to facilitate comparisons between different groups and methodologies, and describe a concrete example of such a task, based on data collected from a spe…

Robust Unsupervised Adaptation of a Speech Recogniser Using Entropy Minimisation and Speaker Codes

2025-06-12 · Rogier C. van Dalen, Shucong Zhang, Titouan Parcollet, Sourav Bhattacharya

Speech recognisers usually perform optimally only in a specific environment and need to be adapted to work well in another. For adaptation to a new speaker, there is often too little data for fine-tuning to be robust, an…

Pseudo Label