paper-with-me

홈 › Papers

NEO-BENCH: Evaluating Robustness of Large Language Models with Neologisms

2024-02-19 · Jonathan Zheng, Alan Ritter, Wei Xu

The performance of Large Language Models (LLMs) degrades from the temporal drift between data used for model training and newer text seen during inference. One understudied avenue of language change causing data drift is the emergence of neologisms -- new word forms -- over time. We create a diverse resource of recent English neologisms by using several popular collection methods. We analyze temporal drift using neologisms by comparing sentences containing new words with near-identical sentences that replace neologisms with existing substitute words. Model performance is nearly halved in machine translation when a single neologism is introduced in a sentence. Motivated by these results, we construct a benchmark to evaluate LLMs' ability to generalize to neologisms with various natural language understanding tasks and model perplexity. Models with later knowledge cutoff dates yield lower perplexities and perform better in downstream tasks. LLMs are also affected differently based on the linguistic origins of words, indicating that neologisms are complex for static LLMs to address. We will release our benchmark and code for reproducing our experiments.

📄 PDF Abstract BibTeX arXiv:2402.12261

Code (1)

jonathanqzheng/neo-bench 공식 구현

Tasks

Machine TranslationNatural Language UnderstandingSentence

Similar Papers 제목 키워드 기반

CNeo-Bench: Diagnosing Large Language Models on Chinese Neologisms

2026-08-28 · Kaiyan Zhao, Zhongtao Miao, Zheyong Xie, Shaosheng Cao 외 arxiv

Chinese neologisms exploit diverse and unique linguistic mechanisms, such as phonetic substitution (e.g., 886 for ``bye-bye'') and visual character decomposition that are rare in other languages. We introduce CNeo-Bench,…

Skill Neologisms: Towards Skill-based Continual Learning

2026-05-06 · Antonin Berthon, Nicolas Astorga, Mihaela van der Schaar arxiv

Modern LLMs show mastery over an ever-growing range of skills, as well as the ability to compose them flexibly. However, extending model capabilities to new skills in a scalable manner is an open problem: fine-tuning and…

Continual Learning

A primer on getting neologisms from foreign languages to under-resourced languages

2023-03-07 · Luis Camacho

Mainly due to lack of support, most under-resourced languages have a reduced lexicon in most realms and domains of increasing importance, then their speakers need to significantly augment it. Although neologisms should a…

Common Sense Reasoning

Unsupervised Neologism Normalization Using Embedding Space Mapping

2019-11-01 · WS 2019 11 · Nasser Zalmout, Kapil Thadani, Aasish Pappu

This paper presents an approach for detecting and normalizing neologisms in social media content. Neologisms refer to recent expressions that are specific to certain entities or events and are being increasingly used by …

Natural Language UnderstandingText Normalization

Classification and Analysis of Neologisms Produced by Learners of Spanish: Effects of Proficiency and Task

2020-07-01 · WS 2020 7 · Shira Wein

The Spanish Learner Language Oral Corpora (SPLLOC) of transcribed conversations between investigators and language learners contains a set of neologism tags. In this work, the utterances tagged as neologisms are broken d…

General ClassificationLanguage Acquisition