paper-with-me

Papers

Grammatical Error Annotation for Korean Learners of Spoken English

2012-05-01 · LREC 2012 5 · Hongsuck Seo, Kyusong Lee, Gary Geunbae Lee, Soo-Ok Kweon, Hae-Ri Kim

The goal of our research is to build a grammatical error-tagged corpus for Korean learners of Spoken English dubbed Postech Learner Corpus. We collected raw story-telling speech from Korean university students. Transcription and annotation using the Cambridge Learner Corpus tagset were performed by six Korean annotators fluent in English. For the annotation of the corpus, we developed an annotation tool and a validation tool. After comparing human annotation with machine-recommended error tags, unmatched errors were rechecked by a native annotator. We observed different characteristics between the spoken language corpus built in this study and an existing written language corpus.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Grammatical Error Detection

Similar Papers 제목 키워드 기반

Towards standardizing Korean Grammatical Error Correction: Datasets and Annotation

2022-10-25 · Soyoung Yoon, Sungjoon Park, Gyuwan Kim, Junhee Cho 외

Research on Korean grammatical error correction (GEC) is limited, compared to other major languages such as English. We attribute this problematic circumstance to the lack of a carefully designed evaluation benchmark for…

AttributeDiversityGrammatical Error Correction

Grammatical error detection in transcriptions of spoken English

2020-12-01 · COLING 2020 8 · Andrew Caines, Christian Bentz, Kate Knill, Marek Rei 외

We describe the collection of transcription corrections and grammatical error annotations for the CrowdED Corpus of spoken English monologues on business topics. The corpus recordings were crowdsourced from native speake…

Grammatical Error CorrectionGrammatical Error Detection

Refining Word-Based Grammatical Error Annotation for L2 Korean

2026-05-28 · Jungyeul Park, Kyungtae Lim, Wonjun Oh, Benjamin Nguyen 외 arxiv

Korean grammatical error correction (K-GEC) presents a structural mismatch between word-based evaluation and the morpheme-level locus of many learner errors. Postpositions and verbal endings are bound to lexical hosts, b…

Grammatical Error Correction

Data Augmentation for Spoken Grammatical Error Correction

2025-07-25 · Penny Karanasou, Mengjie Qian, Stefano Bannò, Mark J. F. Gales 외 arxiv

While there exist strong benchmark datasets for grammatical error correction (GEC), high-quality annotated spoken datasets for Spoken GEC (SGEC) are still under-resourced. In this paper, we propose a fully automated meth…

Grammatical Error CorrectionData Augmentation

Korean Children's Spoken English Corpus and an Analysis of its Pronunciation Variability

2012-05-01 · LREC 2012 5 · Hyejin Hong, Sunhee Kim, Minhwa Chung

This paper introduces a corpus of Korean-accented English speech produced by children (the Korean Children's Spoken English Corpus: the KC-SEC), which is constructed by Seoul National University. The KC-SEC was developed…

Speech Recognition