paper-with-me

Papers

Improving Native Language Identification by Using Spelling Errors

2017-07-01 · ACL 2017 7 · Lingzhen Chen, Carlo Strapparava, Vivi Nastase

In this paper, we explore spelling errors as a source of information for detecting the native language of a writer, a previously under-explored area. We note that character n-grams from misspelled words are very indicative of the native language of the author. In combination with other lexical features, spelling error features lead to 1.2{\%} improvement in accuracy on classifying texts in the TOEFL11 corpus by the author{'}s native language, compared to systems participating in the NLI shared task.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Language IdentificationNative Language Identification

Similar Papers 제목 키워드 기반

Anglicized Words and Misspelled Cognates in Native Language Identification

2019-08-01 · WS 2019 8 · Ilia Markov, Vivi Nastase, Carlo Strapparava

In this paper, we present experiments that estimate the impact of specific lexical choices of people writing in a second language (L2). In particular, we look at misspelled words that indicate lexical uncertainty on the …

Language IdentificationNative Language Identification

Native Language Identification with Large Language Models

2023-12-13 · Wei zhang, Alexandre Salle

We present the first experiments on Native Language Identification (NLI) using LLMs such as GPT-4. NLI is the task of predicting a writer's first language by analyzing their writings in a second language, and is used in …

Language AcquisitionLanguage IdentificationNative Language Identification

Neural Networks and Spelling Features for Native Language Identification

2017-09-01 · WS 2017 9 · Johannes Bjerva, Gintar{\.e} Grigonyt{\.e}, Robert {\"O}stling, Barbara Plank

We present the RUG-SU team{'}s submission at the Native Language Identification Shared Task 2017. We combine several approaches into an ensemble, based on spelling error features, a simple neural network using word repre…

Language IdentificationNative Language IdentificationWord Embeddings

CCTC: A Cross-Sentence Chinese Text Correction Dataset for Native Speakers

2022-10-01 · COLING 2022 10 · Baoxin Wang, Xingyi Duan, Dayong Wu, Wanxiang Che 외

The Chinese text correction (CTC) focuses on detecting and correcting Chinese spelling errors and grammatical errors. Most existing datasets of Chinese spelling check (CSC) and Chinese grammatical error correction (GEC) …

Grammatical Error CorrectionSentence

Similarity-Based Unsupervised Spelling Correction Using BioWordVec: Development and Usability Study of Bacterial Culture and Antimicrobial Susceptibility Reports

2021-02-22 · JMIR Medical Informatics 2021 2 · Taehyeong Kim, Sung Won Han, Minji Kang, Se Ha Lee 외

Background: Existing bacterial culture test results for infectious diseases are written in unrefined text, resulting in many problems, including typographical errors and stop words. Effective spelling correction process…

Cultural Vocal Bursts Intensity PredictionSpelling Correction