Computing Optimal Location of Microphone for Improved Speech Recognition
It was shown in our earlier work that the measurement error in the microphone position affected the room impulse response (RIR) which in turn affected the single-channel close microphone and multi-channel distant microphone speech recognition. In this paper, as an extension, we systematically study to identify the optimal location of the microphone, given an approximate and hence erroneous location of the microphone in 3D space. The primary idea is to use Monte-Carlo technique to generate a large number of random microphone positions around the erroneous microphone position and select the microphone position that results in the best performance of a general purpose automatic speech recognition (gp-asr). We experiment with clean and noisy speech and show that the optimal location of the microphone is unique and is affected by noise.
Code (0)
등록된 구현이 없습니다.
Tasks
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)PositionRoom Impulse Response (RIR)speech-recognitionSpeech RecognitionSimilar Papers 제목 키워드 기반
Impact of Microphone position Measurement Error on Multi Channel Distant Speech Recognition & Intelligibility
It was shown in (Raikar et al., 2020) that the measurement error in the microphone position affected the room impulse response (RIR) which in turn affected the single channel speech recognition. In this paper, we ex-tend…
Distant Speech RecognitionPositionRoom Impulse Response (RIR)speech-recognition+2Multiple Speaker Separation from Noisy Sources in Reverberant Rooms using Relative Transfer Matrix
Separation of simultaneously active multiple speakers is a difficult task in environments with strong reverberation and many background noise sources. This paper uses the relative transfer matrix (ReTM), a generalization…
Speaker SeparationEnd-to-end Microphone Permutation and Number Invariant Multi-channel Speech Separation
An important problem in ad-hoc microphone speech separation is how to guarantee the robustness of a system with respect to the locations and numbers of microphones. The former requires the system to be invariant to diffe…
Speech SeparationDistributed Microphone Speech Enhancement based on Deep Learning
Speech-related applications deliver inferior performance in complex noise environments. Therefore, this study primarily addresses this problem by introducing speech-enhancement (SE) systems based on deep neural networks …
AllDeep LearningSpeech EnhancementEnhanced Deep Speech Separation in Clustered Ad Hoc Distributed Microphone Environments
Ad-hoc distributed microphone environments, where microphone locations and numbers are unpredictable, present a challenge to traditional deep learning models, which typically require fixed architectures. To tailor deep l…
Deep LearningSpeech Separation