paper-with-me

홈 › Papers

Attention model for articulatory features detection

2019-07-02 · Ievgen Karaulov, Dmytro Tkanov

Articulatory distinctive features, as well as phonetic transcription, play important role in speech-related tasks: computer-assisted pronunciation training, text-to-speech conversion (TTS), studying speech production mechanisms, speech recognition for low-resourced languages. End-to-end approaches to speech-related tasks got a lot of traction in recent years. We apply Listen, Attend and Spell~(LAS)~\cite{Chan-LAS2016} architecture to phones recognition on a small small training set, like TIMIT~\cite{TIMIT-1992}. Also, we introduce a novel decoding technique that allows to train manners and places of articulation detectors end-to-end using attention models. We also explore joint phones recognition and articulatory features detection in multitask learning setting.

📄 PDF Abstract BibTeX arXiv:1907.01914

Code (1)

sciforce/phones-las 공식 구현 tf

Tasks

Manner Of Articulation Detectionmodelspeech-recognitionSpeech Recognitiontext-to-speechText to Speech

Similar Papers 제목 키워드 기반

Acoustic-to-Articulatory Speech Inversion Features for Mispronunciation Detection of /r/ in Child Speech Sound Disorders

2023-05-25 · Nina R Benway, Yashish M Siriwardena, Jonathan L Preston, Elaine Hitchcock 외

Acoustic-to-articulatory speech inversion could enhance automated clinical mispronunciation detection to provide detailed articulatory feedback unattainable by formant-based mispronunciation detection algorithms; however…

Articulation-Informed ASR: Integrating Articulatory Features into ASR via Auxiliary Speech Inversion and Cross-Attention Fusion

2025-10-01 · Ahmed Adel Attia, Jing Liu, Carol Espy Wilson arxiv

Prior works have investigated the use of articulatory features as complementary representations for automatic speech recognition (ASR), but their use was largely confined to shallow acoustic models. In this work, we revi…

Speech Recognition

A comparative study of estimating articulatory movements from phoneme sequences and acoustic features

2019-10-31 · Abhayjeet Singh, Aravind Illa, Prasanta Kumar Ghosh

Unlike phoneme sequences, movements of speech articulators (lips, tongue, jaw, velum) and the resultant acoustic signal are known to encode not only the linguistic message but also carry para-linguistic information. Whil…

Advancing Speech Synthesis using EEG

2020-04-09 · Gautam Krishna, Co Tran, Mason Carnahan, Ahmed Tewfik

In this paper we introduce attention-regression model to demonstrate predicting acoustic features from electroencephalography (EEG) features recorded in parallel with spoken sentences. First we demonstrate predicting aco…

EEGElectroencephalogram (EEG)regressionSpeech Synthesis

Improving generalization of vocal tract feature reconstruction: from augmented acoustic inversion to articulatory feature reconstruction without articulatory data

2018-09-04 · Rosanna Turrisi, Raffaele Tavarone, Leonardo Badino

We address the problem of reconstructing articulatory movements, given audio and/or phonetic labels. The scarce availability of multi-speaker articulatory data makes it difficult to learn a reconstruction that generalize…