paper-with-me

홈 › Papers

A Database of Freely Written Texts of German School Students for the Purpose of Automatic Spelling Error Classification

2014-05-01 · LREC 2014 5 · Kay Berkling, Johanna Fay, Masood Ghayoomi, Katrin Hein, R{\'e}mi Lavalley, Ludwig Linhuber, Sebastian St{\"u}ker

The spelling competence of school students is best measured on freely written texts, instead of pre-determined, dictated texts. Since the analysis of the error categories in these kinds of texts is very labor intensive and costly, we are working on an automatic systems to perform this task. The modules of the systems are derived from techniques from the area of natural language processing, and are learning systems that need large amounts of training data. To obtain the data necessary for training and evaluating the resulting system, we conducted data collection of freely written, German texts by school children. 1,730 students from grade 1 through 8 participated in this data collection. The data was transcribed electronically and annotated with their corrected version. This resulted in a total of 14,563 sentences that can now be used for research regarding spelling diagnostics. Additional meta-data was collected regarding writers{'} language biography, teaching methodology, age, gender, and school year. In order to do a detailed manual annotation of the categories of the spelling errors committed by the students we developed a tool specifically tailored to the task.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

General Classification

Similar Papers 제목 키워드 기반

The making of the Litkey Corpus, a richly annotated longitudinal corpus of German texts written by primary school children

2019-08-01 · WS 2019 8 · Ronja Laarmann-Quante, Stefanie Dipper, Eva Belke

To date, corpus and computational linguistic work on written language acquisition has mostly dealt with second language learners who have usually already mastered orthography acquisition in their first language. In this …

Language AcquisitionPOS

CVL-DataBase: An Off-Line Database for Writer Retrieval, Writer Identification and Word Spotting

2013-08-25 · International Conference on Document Analysis and Recognition 2013 8 · Florian Kleber, Stefan Fiel, Markus Diem, Robert Sablatnig

In this paper a public database for writer retrieval, writer identification and word spotting is presented. The CVL-Database consists of 7 different handwritten texts (1 German and 6 English Texts) and 311 different writ…

RetrievalWriter Retrieval

Annotating Spelling Errors in German Texts Produced by Primary School Children

2016-08-01 · WS 2016 8 · Ronja Laarmann-Quante, Lukas Knichel, Stefanie Dipper, Carina Betken

E:Calm Resource: a Resource for Studying Texts Produced by French Pupils and Students

2020-05-01 · LREC 2020 5 · Lydia-Mai Ho-Dac, Serge Fleury, Claude Ponton

The E:Calm resource is constructed from French student texts produced in a variety of usual contexts of teaching. The distinction of the E:Calm resource is to provide an ecological data set that gives a broad overview of…

POSPOS Tagging

TLT-school: a Corpus of Non Native Children Speech

2020-01-22 · LREC 2020 5 · Roberto Gretter, Marco Matassoni, Stefano Bannò, Daniele Falavigna

This paper describes "TLT-school" a corpus of speech utterances collected in schools of northern Italy for assessing the performance of students learning both English and German. The corpus was recorded in the years 2017…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition