paper-with-me

Noisy Speech Recognition

2개 벤치마크 · 논문 13편 · 이 태스크의 논문 보기 →

Benchmarks

CHiME real

결과 5개

CHiME clean

결과 2개

Most implemented

Papers

Efficient Extraction of Noise-Robust Discrete Units from Self-Supervised Speech Models

2024-09-04 · Jakob Poncelet, Yujun Wang, Hugo Van hamme

Continuous speech can be converted into a discrete sequence by deriving discrete units from the hidden features of self-supervised learned (SSL) speech models. Although SSL models are becoming larger and trained on more …

DecoderNoisy Speech Recognitionspeech-recognitionSpeech Recognition

Direction-Aware Joint Adaptation of Neural Speech Enhancement and Recognition in Real Multiparty Conversational Environments

2022-07-15 · Yicheng Du, Aditya Arie Nugraha, Kouhei Sekiguchi, Yoshiaki Bando 외

This paper describes noisy speech recognition for an augmented reality headset that helps verbal communication within real multiparty conversational environments. A major approach that has actively been studied in simula…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Distant Speech RecognitionNoisy Speech Recognition+3

Visual Context-driven Audio Feature Enhancement for Robust End-to-End Audio-Visual Speech Recognition

2022-07-13 · Joanna Hong, Minsu Kim, Daehun Yoo, Yong Man Ro

This paper focuses on designing a noise-robust end-to-end Audio-Visual Speech Recognition (AVSR) system. To this end, we propose Visual Context-driven Audio Feature Enhancement module (V-CAFE) to enhance the input noisy …

Audio-Visual Speech RecognitionDecoderNoisy Speech Recognitionspeech-recognition+2

Improving Noise Robustness of Contrastive Speech Representation Learning with Speech Reconstruction

2021-10-28 · Heming Wang, Yao Qian, Xiaofei Wang, Yiming Wang 외

Noise robustness is essential for deploying automatic speech recognition (ASR) systems in real-world environments. One way to reduce the effect of noise interference is to employ a preprocessing module that conducts spee…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Auxiliary LearningContrastive Learning+8

Speech Recognition With No Speech Or With Noisy Speech Beyond English

2019-06-17 · Gautam Krishna, Co Tran, Yan Han, Mason Carnahan 외

In this paper we demonstrate continuous noisy speech recognition using connectionist temporal classification (CTC) model on limited Chinese vocabulary using electroencephalography (EEG) features with no speech signal as …

EEGElectroencephalogram (EEG)General ClassificationNoisy Speech Recognition+2

An Investigation of End-to-End Multichannel Speech Recognition for Reverberant and Mismatch Conditions

2019-04-19 · Aswin Shanmugam Subramanian, Xiaofei Wang, Shinji Watanabe, Toru Taniguchi 외

Sequence-to-sequence (S2S) modeling is becoming a popular paradigm for automatic speech recognition (ASR) because of its ability to jointly optimize all the conventional ASR components in an end-to-end (E2E) fashion. Thi…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)DenoisingNoisy Speech Recognition+3

전체 13편 보기 →