Flowchase: a Mobile Application for Pronunciation Training
In this paper, we present a solution for providing personalized and instant feedback to English learners through a mobile application, called Flowchase, that is connected to a speech technology able to segment and analyze speech segmental and supra-segmental features. The speech processing pipeline receives linguistic information corresponding to an utterance to analyze along with a speech sample. After validation of the speech sample, a joint forced-alignment and phonetic recognition is performed thanks to a combination of machine learning models based on speech representation learning that provides necessary information for designing a feedback on a series of segmental and supra-segmental pronunciation aspects.
Code (0)
등록된 구현이 없습니다.
Tasks
Representation LearningSpeech Representation LearningSimilar Papers 제목 키워드 기반
Enhancing nonnative speech perception and production through an AI-powered application
While research on using Artificial Intelligence (AI) through various applications to enhance foreign language pronunciation is expanding, it has primarily focused on aspects such as comprehensibility and intelligibility,…
SentenceAnomaly detection with a variational autoencoder for Arabic mispronunciation detection
Computer-assisted language learning (CALL) systems increasingly arouse a significant interest and establish a presence in automated foreign language learning. They enhance traditional learning methods by providing acces…
Anomaly DetectionCompress Polyphone Pronunciation Prediction Model with Shared Labels
It is well known that deep learning model has huge parameters and is computationally expensive, especially for embedded and mobile devices. Polyphone pronunciations selection is a basic function for Chinese Text-to-Speec…
PredictionQuantizationtext-to-speechText to SpeechAutomatic Pronunciation Assessment -- A Review
Pronunciation assessment and its application in computer-aided pronunciation training (CAPT) have seen impressive progress in recent years. With the rapid growth in language processing and deep learning over the past few…
Speaker Independent Continuous Speech to Text Converter for Mobile Application
An efficient speech to text converter for mobile application is presented in this work. The prime motive is to formulate a system which would give optimum performance in terms of complexity, accuracy, delay and memory re…
Action DetectionActivity DetectionSpeech-to-Text