paper-with-me

홈 › Papers

DDSupport: Language Learning Support System that Displays Differences and Distances from Model Speech

2022-12-08 · Kazuki Kawamura, Jun Rekimoto

When beginners learn to speak a non-native language, it is difficult for them to judge for themselves whether they are speaking well. Therefore, computer-assisted pronunciation training systems are used to detect learner mispronunciations. These systems typically compare the user's speech with that of a specific native speaker as a model in units of rhythm, phonemes, or words and calculate the differences. However, they require extensive speech data with detailed annotations or can only compare with one specific native speaker. To overcome these problems, we propose a new language learning support system that calculates speech scores and detects mispronunciations by beginners based on a small amount of unannotated speech data without comparison to a specific person. The proposed system uses deep learning--based speech processing to display the pronunciation score of the learner's speech and the difference/distance between the learner's and a group of models' pronunciation in an intuitively visual manner. Learners can gradually improve their pronunciation by eliminating differences and shortening the distance from the model until they become sufficiently proficient. Furthermore, since the pronunciation score and difference/distance are not calculated compared to specific sentences of a particular model, users are free to study the sentences they wish to study. We also built an application to help non-native speakers learn English and confirmed that it can improve users' speech intelligibility.

📄 PDF Abstract BibTeX arXiv:2212.04930

Code (0)

등록된 구현이 없습니다.

Tasks

Rhythm

Similar Papers 제목 키워드 기반

Liquid-crystal display (LCD) of achromatic, mean-modulated flicker in clinical assessment and experimental studies of visual systems

2020-11-17

Achromatic, mean-modulated flicker (wherein luminance increments and decrements of equal magnitude are applied, over time, to a test field) is commonly used in both clinical assessment of vision and experimental studies …

Learning Word Groundings from Humans Facilitated by Robot Emotional Displays

2020-07-01 · SIGDIAL (ACL) 2020 7 · David McNeill, Casey Kennington

In working towards accomplishing a human-level acquisition and understanding of language, a robot must meet two requirements: the ability to learn words from interactions with its physical environment, and the ability to…

LLM BiasScope: A Real-Time Bias Analysis Platform for Comparative LLM Evaluation

2026-03-12 · Himel Ghosh, Nick Elias Werner arxiv

As large language models (LLMs) are deployed widely, detecting and understanding bias in their outputs is critical. We present LLM BiasScope, a web application for side-by-side comparison of LLM outputs with real-time bi…

Bias Detection

Emergence of a phonological bias in ChatGPT

2023-05-25 · Juan Manuel Toro

Current large language models, such as OpenAI's ChatGPT, have captured the public's attention because how remarkable they are in the use of language. Here, I demonstrate that ChatGPT displays phonological biases that are…

Chatbot

Test Models for Statistical Inference: Two-Dimensional Reaction Systems Displaying Limit Cycle Bifurcations and Bistability

2017-05-29

Theoretical results regarding two-dimensional ordinary-differential equations (ODEs) with second-degree polynomial right-hand sides are summarized, with an emphasis on limit cycles, limit cycle bifurcations and multistab…

Time SeriesTime Series Analysis