paper-with-me

홈 › Papers

FAB: The French Absolute Beginner Corpus for Pronunciation Training

2020-05-01 · LREC 2020 5 · Sean Robertson, Cosmin Munteanu, Gerald Penn

We introduce the French Absolute Beginner (FAB) speech corpus. The corpus is intended for the development and study of Computer-Assisted Pronunciation Training (CAPT) tools for absolute beginner learners. Data were recorded during two experiments focusing on using a CAPT system in paired role-play tasks. The setting grants FAB three distinguishing features from other non-native corpora: the experimental setting is ecologically valid, closing the gap between training and deployment; it features a label set based on teacher feedback, allowing for context-sensitive CAPT; and data have been primarily collected from absolute beginners, a group often ignored. Participants did not read prompts, but instead recalled and modified dialogues that were modelled in videos. Unable to distinguish modelled words solely from viewing videos, speakers often uttered unintelligible or out-of-L2 words. The corpus is split into three partitions: one from an experiment with minimal feedback; another with explicit, word-level feedback; and a third with supplementary read-and-record data. A subset of words in the first partition has been labelled as more or less native, with inter-annotator agreement reported. In the explicit feedback partition, labels are derived from the experiment{'}s online feedback. The FAB corpus is scheduled to be made freely available by the end of 2020.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

valid

Similar Papers 제목 키워드 기반

Neural TTS in French: Comparing Graphemic and Phonetic Inputs Using the SynPaFlex-Corpus and Tacotron2

2023-04-17 · Samuel Delalez, Ludi Akue

The SynPaFlex-Corpus is a publicly available TTS-oriented dataset, which provides phonetic transcriptions automatically produced by the JTrans transcriber, with a Phoneme Error Rate (PER) of 6.1%. In this paper, we analy…

FrSemCor: Annotating a French Corpus with Supersenses

2020-05-01 · LREC 2020 5 · Lucie Barque, Pauline Haas, Richard Huyghe, Delphine Tribout 외

French, as many languages, lacks semantically annotated corpus data. Our aim is to provide the linguistic and NLP research communities with a gold standard sense-annotated corpus of French, using WordNet Unique Beginners…

The IFCASL Corpus of French and German Non-native and Native Read Speech

2016-05-01 · LREC 2016 5 · Juergen Trouvain, Anne Bonneau, Vincent Colotte, Camille Fauth 외

The IFCASL corpus is a French-German bilingual phonetic learner corpus designed, recorded and annotated in a project on individualized feedback in computer-assisted spoken language learning. The motivation for setting up…

DDSupport: Language Learning Support System that Displays Differences and Distances from Model Speech

2022-12-08 · Kazuki Kawamura, Jun Rekimoto

When beginners learn to speak a non-native language, it is difficult for them to judge for themselves whether they are speaking well. Therefore, computer-assisted pronunciation training systems are used to detect learner…

Rhythm

A Two-Step Approach for Data-Efficient French Pronunciation Learning

2024-10-08 · Hoyeon Lee, Hyeeun Jang, Jong-Hwan Kim, Jae-Min Kim

Recent studies have addressed intricate phonological phenomena in French, relying on either extensive linguistic knowledge or a significant amount of sentence-level pronunciation data. However, creating such resources is…

Sentence