paper-with-me

Papers

Complex Word Identification Based on Frequency in a Learner Corpus

2018-06-01 · WS 2018 6 · Tomoyuki Kajiwara, Mamoru Komachi

We introduce the TMU systems for the Complex Word Identification (CWI) Shared Task 2018. TMU systems use random forest classifiers and regressors whose features are the number of characters, the number of words, and the frequency of target words in various corpora. Our simple systems performed best on 5 tracks out of 12 tracks. Our ablation analysis revealed the usefulness of a learner corpus for CWI task.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Complex Word IdentificationLexical SimplificationReading ComprehensionText Simplification

Similar Papers 제목 키워드 기반

CLexIS2: A New Corpus for Complex Word Identification Research in Computing Studies

2021-09-01 · RANLP 2021 9 · Jenny A. Ortiz Zambrano, Arturo Montejo-Ráez

Reading is a complex process not only because of the words or sections that are difficult for the reader to understand. Complex word identification (CWI) is the task of detecting in the content of documents the words tha…

Complex Word IdentificationLexical Simplification

SeCoDa: Sense Complexity Dataset

2020-05-01 · LREC 2020 5 · David Strohmaier, Sian Gooding, Shiva Taslimipoor, Ekaterina Kochmar

The Sense Complexity Dataset (SeCoDa) provides a corpus that is annotated jointly for complexity and word senses. It thus provides a valuable resource for both word sense disambiguation and the task of complex word ident…

Complex Word IdentificationWord Sense Disambiguation

Multi-task Learning for Chinese Word Usage Errors Detection

2019-04-03 · Jinbin Zhang, Heng Wang

Chinese word usage errors often occur in non-native Chinese learners' writing. It is very helpful for non-native Chinese learners to detect them automatically when learning writing. In this paper, we propose a novel appr…

Multi-Task LearningPOSPOS TaggingPrediction

Complex Word Identification in Vietnamese: Towards Vietnamese Text Simplification

2022-07-01 · NAACL (MIA) 2022 7 · Phuong Nguyen, David Kauchak

Text Simplification has been an extensively researched problem in English, but has not been investigated in Vietnamese. We focus on the Vietnamese-specific Complex Word Identification task, often the first step in Lexica…

Complex Word IdentificationLexical SimplificationText SimplificationVietnamese Datasets

An evaluation of the role of statistical measures and frequency for MWE identification

2014-05-01 · LREC 2014 5 · S Antunes, ra, Am{\'a}lia Mendes

We report on an experiment to evaluate the role of statistical association measures and frequency for the identification of MWE. We base our evaluation on a lexicon of 14.000 MWE comprising different types of word combin…