paper-with-me

홈 › Papers

Revisiting Grammatical Error Correction Evaluation and Beyond

2022-11-03 · Peiyuan Gong, Xuebo Liu, Heyan Huang, Min Zhang

Pretraining-based (PT-based) automatic evaluation metrics (e.g., BERTScore and BARTScore) have been widely used in several sentence generation tasks (e.g., machine translation and text summarization) due to their better correlation with human judgments over traditional overlap-based methods. Although PT-based methods have become the de facto standard for training grammatical error correction (GEC) systems, GEC evaluation still does not benefit from pretrained knowledge. This paper takes the first step towards understanding and improving GEC evaluation with pretraining. We first find that arbitrarily applying PT-based metrics to GEC evaluation brings unsatisfactory correlation results because of the excessive attention to inessential systems outputs (e.g., unchanged parts). To alleviate the limitation, we propose a novel GEC evaluation metric to achieve the best of both worlds, namely PT-M2 which only uses PT-based metrics to score those corrected parts. Experimental results on the CoNLL14 evaluation task show that PT-M2 significantly outperforms existing methods, achieving a new state-of-the-art result of 0.949 Pearson correlation. Further analysis reveals that PT-M2 is robust to evaluate competitive GEC systems. Source code and scripts are freely available at https://github.com/pygongnlp/PT-M2.

📄 PDF Abstract BibTeX arXiv:2211.01635

Code (1)

pygongnlp/pt-m2 공식 구현 pytorch

Tasks

Grammatical Error CorrectionMachine TranslationSentenceText Summarization

Similar Papers 제목 키워드 기반

Revisiting Meta-evaluation for Grammatical Error Correction

2024-03-05 · Masamune Kobayashi, Masato Mita, Mamoru Komachi

Metrics are the foundation for automatic evaluation in grammatical error correction (GEC), with their evaluation of the metrics (meta-evaluation) relying on their correlation with human judgments. However, conventional m…

Grammatical Error CorrectionSentence

ChatGPT or Grammarly? Evaluating ChatGPT on Grammatical Error Correction Benchmark

2023-03-15 · Haoran Wu, Wenxuan Wang, Yuxuan Wan, Wenxiang Jiao 외

ChatGPT is a cutting-edge artificial intelligence language model developed by OpenAI, which has attracted a lot of attention due to its surprisingly strong ability in answering follow-up questions. In this report, we aim…

Grammatical Error CorrectionLanguage ModelingLanguage ModellingSentence

Evaluating the Capability of Large-scale Language Models on Chinese Grammatical Error Correction Task

2023-07-08 · Fanyi Qu, Yunfang Wu

Large-scale language models (LLMs) has shown remarkable capability in various of Natural Language Processing (NLP) tasks and attracted lots of attention recently. However, some studies indicated that large language model…

Grammatical Error Correction

Towards Automated Document Revision: Grammatical Error Correction, Fluency Edits, and Beyond

2022-05-23 · Masato Mita, Keisuke Sakaguchi, Masato Hagiwara, Tomoya Mizumoto 외

Natural language processing technology has rapidly improved automated grammatical error correction tasks, and the community begins to explore document-level revision as one of the next challenges. To go beyond sentence-l…

Grammatical Error CorrectionLanguage ModellingSentence

MixEdit: Revisiting Data Augmentation and Beyond for Grammatical Error Correction

2023-10-18 · Jingheng Ye, Yinghui Li, Yangning Li, Hai-Tao Zheng

Data Augmentation through generating pseudo data has been proven effective in mitigating the challenge of data scarcity in the field of Grammatical Error Correction (GEC). Various augmentation strategies have been widely…

Data AugmentationDiversityGrammatical Error Correction