paper-with-me

Papers

Acoustic Model Adaptation from Raw Waveforms with SincNet

2019-09-30 · Joachim Fainberg, Ondřej Klejch, Erfan Loweimi, Peter Bell, Steve Renals

Raw waveform acoustic modelling has recently gained interest due to neural networks' ability to learn feature extraction, and the potential for finding better representations for a given scenario than hand-crafted features. SincNet has been proposed to reduce the number of parameters required in raw-waveform modelling, by restricting the filter functions, rather than having to learn every tap of each filter. We study the adaptation of the SincNet filter parameters from adults' to children's speech, and show that the parameterisation of the SincNet layer is well suited for adaptation in practice: we can efficiently adapt with a very small number of parameters, producing error rates comparable to techniques using orders of magnitude more parameters.

📄 PDF Abstract BibTeX arXiv:1909.13759

Code (1)

jfainberg/sincnet_adapt 공식 구현 tf

Tasks

Acoustic Modelling

Similar Papers 제목 키워드 기반

End-to-End Mispronunciation Detection and Diagnosis From Raw Waveforms

2021-03-04 · Bi-Cheng Yan, Berlin Chen

Mispronunciation detection and diagnosis (MDD) is designed to identify pronunciation errors and provide instructive feedback to guide non-native language learners, which is a core component in computer-assisted pronuncia…

Deep Learning For Prominence Detection In Children's Read Speech

2021-10-27 · Mithilesh Vaidya, Kamini Sabu, Preeti Rao

The detection of perceived prominence in speech has attracted approaches ranging from the design of linguistic knowledge-based acoustic features to the automatic feature learning from suprasegmental attributes such as pi…

Deep Learning

Speech recognition for air traffic control via feature learning and end-to-end training

2021-11-04 · Peng Fan, Dongyue Guo, Yi Lin, Bo Yang 외

In this work, we propose a new automatic speech recognition (ASR) system based on feature learning and an end-to-end training procedure for air traffic control (ATC) systems. The proposed model integrates the feature lea…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Speaker Recognition from Raw Waveform with SincNet

2018-07-29 · Mirco Ravanelli, Yoshua Bengio

Deep learning is progressively gaining popularity as a viable alternative to i-vectors for speaker recognition. Promising results have been recently obtained with Convolutional Neural Networks (CNNs) when fed by raw spee…

Speaker IdentificationSpeaker RecognitionSpeaker Verification

PF-Net: Personalized Filter for Speaker Recognition from Raw Waveform

2021-05-31 · Wencheng Li, Zhenhua Tan, Jingyu Ning, Zhenche Xia 외

Speaker recognition using i-vector has been replaced by speaker recognition using deep learning. Speaker recognition based on Convolutional Neural Networks (CNNs) has been widely used in recent years, which learn low-lev…

Speaker IdentificationSpeaker Recognition