paper-with-me

Papers

Automatically Detecting Reduced-formed English Pronunciations by Using Deep Learning

2022-07-01 · NAACL (BEA) 2022 7 · Lei Chen, Chenglin Jiang, Yiwei Gu, Yang Liu, Jiahong Yuan

Reduced form pronunciations are widely used by native English speakers, especially in casual conversations. Second language (L2) learners have difficulty in processing reduced form pronunciations in listening comprehension and face challenges in production too. Meanwhile, training applications dedicated to reduced forms are still few. To solve this issue, we report on our first effort of using deep learning to evaluate L2 learners’ reduced form pronunciations. Compared with a baseline solution that uses an ASR to determine regular or reduced-formed pronunciations, a classifier that learns representative features via a convolution neural network (CNN) on low-level acoustic features, yields higher detection performance. F-1 metric has been increased from $0.690$ to $0.757$ on the reduction task. Furthermore, adding word entities to compute attention weights to better adjust the features learned by the CNN model helps increasing F-1 to $0.763$.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Deep LearningForm

Similar Papers 제목 키워드 기반

Mispronunciation Detection in Non-native (L2) English with Uncertainty Modeling

2021-01-16 · Daniel Korzekwa, Jaime Lorenzo-Trueba, Szymon Zaporowski, Shira Calamaro 외

A common approach to the automatic detection of mispronunciation in language learning is to recognize the phonemes produced by a student and compare it to the expected pronunciation of a native speaker. This approach mak…

Automatic Phoneme RecognitionPhoneme RecognitionSentencevalid

Weakly-supervised word-level pronunciation error detection in non-native English speech

2021-06-07 · Daniel Korzekwa, Jaime Lorenzo-Trueba, Thomas Drugman, Shira Calamaro 외

We propose a weakly-supervised model for word-level mispronunciation detection in non-native (L2) English speech. To train this model, phonetically transcribed L2 speech is not required and we only need to mark mispronou…

Acoustic data-driven lexicon learning based on a greedy pronunciation selection framework

2017-06-12 · Xiaohui Zhang, Vimal Manohar, Daniel Povey, Sanjeev Khudanpur

Speech recognition systems for irregularly-spelled languages like English normally require hand-written pronunciations. In this paper, we describe a system for automatically obtaining pronunciations of words for which pr…

speech-recognitionSpeech Recognition

On Pronunciations in Wiktionary: Extraction and Experiments on Multilingual Syllabification and Stress Prediction

2021-09-01 · RANLP (BUCC) 2021 9 · Winston Wu, David Yarowsky

We constructed parsers for five non-English editions of Wiktionary, which combined with pronunciations from the English edition, comprises over 5.3 million IPA pronunciations, the largest pronunciation lexicon of its kin…

An Investigation of Indian Native Language Phonemic Influences on L2 English Pronunciations

2022-12-19 · Shelly Jain, Priyanshi Pal, Anil Vuppala, Prasanta Ghosh 외

Speech systems are sensitive to accent variations. This is especially challenging in the Indian context, with an abundance of languages but a dearth of linguistic studies characterising pronunciation variations. The grow…