Papers Complex Word Identification
“Complex Word Identification” 태그가 달린 논문 67편 · 필터 해제
New Evaluation Paradigm for Lexical Simplification
Lexical Simplification (LS) methods use a three-step pipeline: complex word identification, substitute generation, and substitute ranking, each with separate evaluation datasets. We found large language models (LLMs) can…
Complex Word IdentificationIn-Context LearningLexical SimplificationSentenceInvestigating Large Language Models for Complex Word Identification in Multilingual and Multidomain Setups
Complex Word Identification (CWI) is an essential step in the lexical simplification task and has recently become a task on its own. Some variations of this binary classification task have emerged, such as lexical comple…
Binary ClassificationComplex Word IdentificationLexical Complexity PredictionLexical Simplification+2Difficult for Whom? A Study of Japanese Lexical Complexity
The tasks of lexical complexity prediction (LCP) and complex word identification (CWI) commonly presuppose that difficult to understand words are shared by the target population. Meanwhile, personalization methods have a…
Complex Word IdentificationLexical Complexity PredictionA User-Centered Evaluation of Spanish Text Simplification
We present an evaluation of text simplification (TS) in Spanish for a production system, by means of two corpora focused in both complex-sentence and complex-word identification. We compare the most prevalent Spanish-spe…
Complex Word IdentificationSentenceText SimplificationComplex Word Identification in Vietnamese: Towards Vietnamese Text Simplification
Text Simplification has been an extensively researched problem in English, but has not been investigated in Vietnamese. We focus on the Vietnamese-specific Complex Word Identification task, often the first step in Lexica…
Complex Word IdentificationLexical SimplificationText SimplificationVietnamese DatasetsCWID-hi: A Dataset for Complex Word Identification in Hindi Text
Text simplification is a method for improving the accessibility of text by converting complex sentences into simple sentences. Multiple studies have been done to create datasets for text simplification. However, most of …
Complex Word IdentificationText SimplificationDomain Adaptation in Multilingual and Multi-Domain Monolingual Settings for Complex Word Identification
Complex word identification (CWI) is a cornerstone process towards proper text simplification. CWI is highly dependent on context, whereas its difficulty is augmented by the scarcity of available datasets which vary grea…
Complex Word IdentificationDomain AdaptationLexical Complexity PredictionText SimplificationOne Size Does Not Fit All: The Case for Personalised Word Complexity Models
Complex Word Identification (CWI) aims to detect words within a text that a reader may find difficult to understand. It has been shown that CWI systems can improve text simplification, readability prediction and vocabula…
Active LearningAllComplex Word IdentificationText SimplificationA Survey on Using Gaze Behaviour for Natural Language Processing
Gaze behaviour has been used as a way to gather cognitive information for a number of years. In this paper, we discuss the use of gaze behaviour in solving different tasks in natural language processing (NLP) without hav…
Complex Word IdentificationSurveyCLexIS2: A New Corpus for Complex Word Identification Research in Computing Studies
Reading is a complex process not only because of the words or sections that are difficult for the reader to understand. Complex word identification (CWI) is the task of detecting in the content of documents the words tha…
Complex Word IdentificationLexical SimplificationIAPUCP at SemEval-2021 Task 1: Stacking Fine-Tuned Transformers is Almost All You Need for Lexical Complexity Prediction
This paper describes our submission to SemEval-2021 Task 1: predicting the complexity score for single words. Our model leverages standard morphosyntactic and frequency-based features that proved helpful for Complex Word…
AllComplex Word IdentificationLexical Complexity PredictionMulti-Task Learning+1Manchester Metropolitan at SemEval-2021 Task 1: Convolutional Networks for Complex Word Identification
We present two convolutional neural networks for predicting the complexity of words and phrases in context on a continuous scale. Both models utilize word and character embeddings alongside lexical features as inputs. Ou…
Complex Word Identificationregressioncs60075_team2 at SemEval-2021 Task 1 : Lexical Complexity Prediction using Transformer-based Language Models pre-trained on various text corpora
This paper describes the performance of the team cs60075_team2 at SemEval 2021 Task 1 - Lexical Complexity Prediction. The main contribution of this paper is to fine-tune transformer-based language models pre-trained on …
Complex Word IdentificationLexical AnalysisLexical Complexity PredictionTask 2Predicting Lexical Complexity in English Texts: The Complex 2.0 Dataset
Identifying words which may cause difficulty for a reader is an essential step in most lexical text simplification systems prior to lexical substitution and can also be used for assessing the readability of a text. This …
Complex Word IdentificationLexical Complexity PredictionText SimplificationCross-Lingual Transfer Learning for Complex Word Identification
Complex Word Identification (CWI) is a task centered on detecting hard-to-understand words, or groups of words, in texts from different areas of expertise. The purpose of CWI is to highlight problematic structures that n…
Complex Word IdentificationCross-Lingual TransferFew-Shot LearningTransfer Learning+1Interpreting Neural CWI Classifiers' Weights as Vocabulary Size
Complex Word Identification (CWI) is a task for the identification of words that are challenging for second-language learners to read. Even though the use of neural classifiers is now common in CWI, the interpretation of…
Complex Word IdentificationDetecting Multiword Expression Type Helps Lexical Complexity Assessment
Multiword expressions (MWEs) represent lexemes that should be treated as single lexical units due to their idiosyncratic nature. Multiple NLP applications have been shown to benefit from MWE identification, however the r…
Complex Word IdentificationText SimplificationVocal Bursts Type PredictionCompLex --- A New Corpus for Lexical Complexity Prediction from Likert Scale Data
Predicting which words are considered hard to understand for a given target population is a vital step in many NLP applications such astext simplification. This task is commonly referred to as Complex Word Identification…
Binary ClassificationComplex Word IdentificationLexical Complexity PredictionSeCoDa: Sense Complexity Dataset
The Sense Complexity Dataset (SeCoDa) provides a corpus that is annotated jointly for complexity and word senses. It thus provides a valuable resource for both word sense disambiguation and the task of complex word ident…
Complex Word IdentificationWord Sense DisambiguationCompLex: A New Corpus for Lexical Complexity Prediction from Likert Scale Data
Predicting which words are considered hard to understand for a given target population is a vital step in many NLP applications such as text simplification. This task is commonly referred to as Complex Word Identificatio…
Binary ClassificationComplex Word IdentificationLexical Complexity PredictionText Simplification