paper-with-me

홈 › Papers

On the Persistent Effects of Lexicality in Large Language Models

2026-06-01 · Hammad Rizwan, Muhammad Umair Haider, Nishant Subramani, Mona T. Diab, A. B. Siddique, Hassan Sajjad arxiv

Representations extracted from large language models (LLMs) play an important role in many downstream applications. However, the structure of these representations is often influenced by lexical overlap rather than semantic content. Our understanding of the relationship between this lexical influence and semantic content, and its implications for downstream tasks, remains limited. In this work, we investigate representations to quantify the effect of lexical overlap relative to semantic content. We consider several adversarial semantic stress tests and further connect our findings to the information theory perspective. We find that lexical influence extends across the depth of models, consistently across architectures, training regimes, and objective functions, including the models trained for semantic similarity. Moreover, we observe a mid-depth region in which both lexical and semantic signals degrade simultaneously, indicating a transitional regime where representations are poor for both surface form and meaning. We further demonstrate the effect of lexical influence on downstream uses of LLMs using summarization and model editing as a case study.

📄 PDF Abstract BibTeX arXiv:2606.02750

Code (0)

등록된 구현이 없습니다.

Tasks

Semantic Similarity

Similar Papers 제목 키워드 기반

Multi-word Lexical Units Recognition in WordNet

2022-06-01 · LREC (MWE) 2022 6 · Marek Maziarz, Ewa Rudnicka, Łukasz Grabowski

WordNet is a state-of-the-art lexical resource used in many tasks in Natural Language Processing, also in multi-word expression (MWE) recognition. However, not all MWEs recorded in WordNet could be indisputably called le…

Allregression

Learning 3D Persistent Embodied World Models

2025-05-05 · Siyuan Zhou, Yilun Du, Yuncong Yang, Lei Han 외

The ability to simulate the effects of future actions on the world is a crucial ability of intelligent embodied agents, enabling agents to anticipate the effects of their actions and make plans accordingly. While a large…

Beyond Means: Topological Causal Effects under Persistent-Homology Ignorability

2026-03-15 · Amir Saki, Usef Faghihi arxiv

Average treatment effects (ATE) and conditional average treatment effects (CATE) are foundational causal estimands, but they target changes in expected outcomes and can miss treatment-induced changes in the shape of outc…

Persistent Instability in LLM's Personality Measurements: Effects of Scale, Reasoning, and Conversation History

2025-08-06 · Tommaso Tosato, Saskia Helbling, Yorguin-Jose Mantilla-Ramos, Mahmood Hegazy 외 arxiv

Large language models require consistent behavioral patterns for safe deployment, yet there are indications of large variability that may lead to an instable expression of personality traits in these models. We present P…

When Errors Become Memories: Causal Pathway Tracing in Multi-Turn Memory-Augmented LLMs

2026-08-31 · Shuyao Xiao, Shengling Wang, Xuan Chen, Ke Chao 외 arxiv

Long-term memory enables large language models (LLMs) to preserve and reuse information across interactions, but it can also turn localized errors into persistent risks. Existing work mainly evaluates whether memory syst…