paper-with-me

Papers

Semi-automatically Annotated Learner Corpus for Russian

2022-06-01 · LREC 2022 6 · Anisia Katinskaia, Maria Lebedeva, Jue Hou, Roman Yangarber

We present ReLCo— the Revita Learner Corpus—a new semi-automatically annotated learner corpus for Russian. The corpus was collected while several thousand L2 learners were performing exercises using the Revita language-learning system. All errors were detected automatically by the system and annotated by type. Part of the corpus was annotated manually—this part was created for further experiments on automatic assessment of grammatical correctness. The Learner Corpus provides valuable data for studying patterns of grammatical errors, experimenting with grammatical error detection and grammatical error correction, and developing new exercises for language learners. Automating the collection and annotation makes the process of building the learner corpus much cheaper and faster, in contrast to the traditional approach of building learner corpora. We make the data publicly available.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Grammatical Error CorrectionGrammatical Error Detection

Similar Papers 제목 키워드 기반

Toward a Paradigm Shift in Collection of Learner Corpora

2020-05-01 · LREC 2020 5 · Anisia Katinskaia, Sardana Ivanova, Roman Yangarber

We present the first version of the longitudinal Revita Learner Corpus (ReLCo), for Russian. In contrast to traditional learner corpora, ReLCo is collected and annotated fully automatically, while students perform exerci…

Russian Error-Annotated Learner English Corpus: a Tool for Computer-Assisted Language Learning

2014-11-01 · WS 2014 11 · Elizaveta Kuzmenko, Andrey Kutuzov

Creating a Corpus for Russian Data-to-Text Generation Using Neural Machine Translation and Post-Editing

2019-08-01 · WS 2019 8 · Anastasia Shimorina, Elena Khasanova, Claire Gardent

In this paper, we propose an approach for semi-automatically creating a data-to-text (D2T) corpus for Russian that can be used to learn a D2T natural language generation model. An error analysis of the output of an Engli…

Data-to-Text GenerationMachine TranslationText GenerationTranslation

Automatically Ranked Russian Paraphrase Corpus for Text Generation

2020-06-17 · WS 2020 7 · Vadim Gudkov, Olga Mitrofanova, Elizaveta Filippskikh

The article is focused on automatic development and ranking of a large corpus for Russian paraphrase generation which proves to be the first corpus of such type in Russian computational linguistics. Existing manually ann…

Paraphrase GenerationSentenceSentence SimilarityText Generation

Automatic Classification of Russian Learner Errors

2022-06-01 · LREC 2022 6 · Alla Rozovskaya

Grammatical Error Correction systems are typically evaluated overall, without taking into consideration performance on individual error types because system output is not annotated with respect to error type. We introduc…

ClassificationGrammatical Error Correction