paper-with-me

Papers

Construction of an Evaluation Corpus for Grammatical Error Correction for Learners of Japanese as a Second Language

2020-05-01 · LREC 2020 5 · Aomi Koyama, Tomoshige Kiyuna, Kenji Kobayashi, Mio Arai, Mamoru Komachi

The NAIST Lang-8 Learner Corpora (Lang-8 corpus) is one of the largest second-language learner corpora. The Lang-8 corpus is suitable as a training dataset for machine translation-based grammatical error correction systems. However, it is not suitable as an evaluation dataset because the corrected sentences sometimes include inappropriate sentences. Therefore, we created and released an evaluation corpus for correcting grammatical errors made by learners of Japanese as a Second Language (JSL). As our corpus has less noise and its annotation scheme reflects the characteristics of the dataset, it is ideal as an evaluation corpus for correcting grammatical errors in sentences written by JSL learners. In addition, we applied neural machine translation (NMT) and statistical machine translation (SMT) techniques to correct the grammar of the JSL learners{'} sentences and evaluated their results using our corpus. We also compared the performance of the NMT system with that of the SMT system.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Grammatical Error CorrectionMachine TranslationNMTTranslation

Similar Papers 제목 키워드 기반

A Simple Yet Effective Corpus Construction Framework for Indonesian Grammatical Error Correction

2024-10-28 · Nankai Lin, Meiyu Zeng, Wentao Huang, Shengyi Jiang 외

Currently, the majority of research in grammatical error correction (GEC) is concentrated on universal languages, such as English and Chinese. Many low-resource languages lack accessible evaluation corpora. How to effici…

Grammatical Error Correction

Cool English: a Grammatical Error Correction System Based on Large Learner Corpora

2018-08-01 · COLING 2018 8 · Yu-Chun Lo, Jhih-Jie Chen, Ching-Yu Yang, Jason Chang

This paper presents a grammatical error correction (GEC) system that provides corrective feedback for essays. We apply the sequence-to-sequence model, which is frequently used in machine translation and text summarizatio…

Grammatical Error CorrectionMachine TranslationText SummarizationTranslation

Automatic Extraction of English Grammar Pattern Correction Rules

2021-10-01 · ROCLING 2021 10 · Kuan-Yu Shen, Yi-Chien Lin, Jason S. Chang

We introduce a method for generating error-correction rules for grammar pattern errors in a given annotated learner corpus. In our approach, annotated edits in the learner corpus are converted into edit rules for correct…

Refining Word-Based Grammatical Error Annotation for L2 Korean

2026-05-28 · Jungyeul Park, Kyungtae Lim, Wonjun Oh, Benjamin Nguyen 외 arxiv

Korean grammatical error correction (K-GEC) presents a structural mismatch between word-based evaluation and the morpheme-level locus of many learner errors. Postpositions and verbal endings are bound to lexical hosts, b…

Grammatical Error Correction

Cross-Corpora Evaluation and Analysis of Grammatical Error Correction Models --- Is Single-Corpus Evaluation Enough?

2019-04-05 · NAACL 2019 6 · Masato Mita, Tomoya Mizumoto, Masahiro Kaneko, Ryo Nagata 외

This study explores the necessity of performing cross-corpora evaluation for grammatical error correction (GEC) models. GEC models have been previously evaluated based on a single commonly applied corpus: the CoNLL-2014 …

Grammatical Error CorrectionNMT