paper-with-me

Papers Complex Word Identification

“Complex Word Identification” 태그가 달린 논문 67편 · 필터 해제

New Evaluation Paradigm for Lexical Simplification

2025-01-25 · Jipeng Qiang, Minjiang Huang, Yi Zhu, Yunhao Yuan 외

Lexical Simplification (LS) methods use a three-step pipeline: complex word identification, substitute generation, and substitute ranking, each with separate evaluation datasets. We found large language models (LLMs) can…

Complex Word IdentificationIn-Context LearningLexical SimplificationSentence

Investigating Large Language Models for Complex Word Identification in Multilingual and Multidomain Setups

2024-11-03 · Răzvan-Alexandru Smădu, David-Gabriel Ion, Dumitru-Clementin Cercel, Florin Pop 외

Complex Word Identification (CWI) is an essential step in the lexical simplification task and has recently become a task on its own. Some variations of this binary classification task have emerged, such as lexical comple…

Binary ClassificationComplex Word IdentificationLexical Complexity PredictionLexical Simplification+2

Difficult for Whom? A Study of Japanese Lexical Complexity

2024-10-24 · Adam Nohejl, Akio Hayakawa, Yusuke Ide, Taro Watanabe

The tasks of lexical complexity prediction (LCP) and complex word identification (CWI) commonly presuppose that difficult to understand words are shared by the target population. Meanwhile, personalization methods have a…

Complex Word IdentificationLexical Complexity Prediction

A User-Centered Evaluation of Spanish Text Simplification

2023-08-15 · Adrian de Wynter, Anthony Hevia, Si-Qing Chen

We present an evaluation of text simplification (TS) in Spanish for a production system, by means of two corpora focused in both complex-sentence and complex-word identification. We compare the most prevalent Spanish-spe…

Complex Word IdentificationSentenceText Simplification

Complex Word Identification in Vietnamese: Towards Vietnamese Text Simplification

2022-07-01 · NAACL (MIA) 2022 7 · Phuong Nguyen, David Kauchak

Text Simplification has been an extensively researched problem in English, but has not been investigated in Vietnamese. We focus on the Vietnamese-specific Complex Word Identification task, often the first step in Lexica…

Complex Word IdentificationLexical SimplificationText SimplificationVietnamese Datasets

CWID-hi: A Dataset for Complex Word Identification in Hindi Text

2022-06-01 · LREC 2022 6 · Gayatri Venugopal, Dhanya Pramod, Ravi Shekhar

Text simplification is a method for improving the accessibility of text by converting complex sentences into simple sentences. Multiple studies have been done to create datasets for text simplification. However, most of …

Complex Word IdentificationText Simplification

Domain Adaptation in Multilingual and Multi-Domain Monolingual Settings for Complex Word Identification

2022-05-15 · ACL 2022 5 · George-Eduard Zaharia, Răzvan-Alexandru Smădu, Dumitru-Clementin Cercel, Mihai Dascalu

Complex word identification (CWI) is a cornerstone process towards proper text simplification. CWI is highly dependent on context, whereas its difficulty is augmented by the scarcity of available datasets which vary grea…

Complex Word IdentificationDomain AdaptationLexical Complexity PredictionText Simplification

One Size Does Not Fit All: The Case for Personalised Word Complexity Models

2022-05-05 · Findings (NAACL) 2022 7 · Sian Gooding, Manuel Tragut

Complex Word Identification (CWI) aims to detect words within a text that a reader may find difficult to understand. It has been shown that CWI systems can improve text simplification, readability prediction and vocabula…

Active LearningAllComplex Word IdentificationText Simplification

A Survey on Using Gaze Behaviour for Natural Language Processing

2021-12-21 · Sandeep Mathias, Diptesh Kanojia, Abhijit Mishra, Pushpak Bhattacharyya

Gaze behaviour has been used as a way to gather cognitive information for a number of years. In this paper, we discuss the use of gaze behaviour in solving different tasks in natural language processing (NLP) without hav…

Complex Word IdentificationSurvey

CLexIS2: A New Corpus for Complex Word Identification Research in Computing Studies

2021-09-01 · RANLP 2021 9 · Jenny A. Ortiz Zambrano, Arturo Montejo-Ráez

Reading is a complex process not only because of the words or sections that are difficult for the reader to understand. Complex word identification (CWI) is the task of detecting in the content of documents the words tha…

Complex Word IdentificationLexical Simplification

IAPUCP at SemEval-2021 Task 1: Stacking Fine-Tuned Transformers is Almost All You Need for Lexical Complexity Prediction

2021-08-01 · SEMEVAL 2021 · Kervy Rivas Rojas, Fernando Alva-Manchego

This paper describes our submission to SemEval-2021 Task 1: predicting the complexity score for single words. Our model leverages standard morphosyntactic and frequency-based features that proved helpful for Complex Word…

AllComplex Word IdentificationLexical Complexity PredictionMulti-Task Learning+1

Manchester Metropolitan at SemEval-2021 Task 1: Convolutional Networks for Complex Word Identification

2021-08-01 · SEMEVAL 2021 · Robert Flynn, Matthew Shardlow

We present two convolutional neural networks for predicting the complexity of words and phrases in context on a continuous scale. Both models utilize word and character embeddings alongside lexical features as inputs. Ou…

Complex Word Identificationregression

cs60075_team2 at SemEval-2021 Task 1 : Lexical Complexity Prediction using Transformer-based Language Models pre-trained on various text corpora

2021-06-04 · Abhilash Nandy, Sayantan Adak, Tanurima Halder, Sai Mahesh Pokala

This paper describes the performance of the team cs60075_team2 at SemEval 2021 Task 1 - Lexical Complexity Prediction. The main contribution of this paper is to fine-tune transformer-based language models pre-trained on …

Complex Word IdentificationLexical AnalysisLexical Complexity PredictionTask 2

Predicting Lexical Complexity in English Texts: The Complex 2.0 Dataset

2021-02-17 · Matthew Shardlow, Richard Evans, Marcos Zampieri

Identifying words which may cause difficulty for a reader is an essential step in most lexical text simplification systems prior to lexical substitution and can also be used for assessing the readability of a text. This …

Complex Word IdentificationLexical Complexity PredictionText Simplification

Cross-Lingual Transfer Learning for Complex Word Identification

2020-10-02 · George-Eduard Zaharia, Dumitru-Clementin Cercel, Mihai Dascalu

Complex Word Identification (CWI) is a task centered on detecting hard-to-understand words, or groups of words, in texts from different areas of expertise. The purpose of CWI is to highlight problematic structures that n…

Complex Word IdentificationCross-Lingual TransferFew-Shot LearningTransfer Learning+1

Interpreting Neural CWI Classifiers' Weights as Vocabulary Size

2020-07-01 · WS 2020 7 · Yo Ehara

Complex Word Identification (CWI) is a task for the identification of words that are challenging for second-language learners to read. Even though the use of neural classifiers is now common in CWI, the interpretation of…

Complex Word Identification

Detecting Multiword Expression Type Helps Lexical Complexity Assessment

2020-05-12 · LREC 2020 5 · Ekaterina Kochmar, Sian Gooding, Matthew Shardlow

Multiword expressions (MWEs) represent lexemes that should be treated as single lexical units due to their idiosyncratic nature. Multiple NLP applications have been shown to benefit from MWE identification, however the r…

Complex Word IdentificationText SimplificationVocal Bursts Type Prediction

CompLex --- A New Corpus for Lexical Complexity Prediction from Likert Scale Data

2020-05-01 · LREC 2020 5 · Matthew Shardlow, Michael Cooper, Marcos Zampieri

Predicting which words are considered hard to understand for a given target population is a vital step in many NLP applications such astext simplification. This task is commonly referred to as Complex Word Identification…

Binary ClassificationComplex Word IdentificationLexical Complexity Prediction

SeCoDa: Sense Complexity Dataset

2020-05-01 · LREC 2020 5 · David Strohmaier, Sian Gooding, Shiva Taslimipoor, Ekaterina Kochmar

The Sense Complexity Dataset (SeCoDa) provides a corpus that is annotated jointly for complexity and word senses. It thus provides a valuable resource for both word sense disambiguation and the task of complex word ident…

Complex Word IdentificationWord Sense Disambiguation

CompLex: A New Corpus for Lexical Complexity Prediction from Likert Scale Data

2020-03-16 · Matthew Shardlow, Michael Cooper, Marcos Zampieri

Predicting which words are considered hard to understand for a given target population is a vital step in many NLP applications such as text simplification. This task is commonly referred to as Complex Word Identificatio…

Binary ClassificationComplex Word IdentificationLexical Complexity PredictionText Simplification
1–20 / 67 다음 →