paper-with-me

홈 › Papers

Limits for Learning with Language Models

2023-06-21 · Nicholas Asher, Swarnadeep Bhar, Akshay Chaturvedi, Julie Hunter, Soumya Paul

With the advent of large language models (LLMs), the trend in NLP has been to train LLMs on vast amounts of data to solve diverse language understanding and generation tasks. The list of LLM successes is long and varied. Nevertheless, several recent papers provide empirical evidence that LLMs fail to capture important aspects of linguistic meaning. Focusing on universal quantification, we provide a theoretical foundation for these empirical findings by proving that LLMs cannot learn certain fundamental semantic properties including semantic entailment and consistency as they are defined in formal semantics. More generally, we show that LLMs are unable to learn concepts beyond the first level of the Borel Hierarchy, which imposes severe limits on the ability of LMs, both large and small, to capture many aspects of linguistic meaning. This means that LLMs will continue to operate without formal guarantees on tasks that require entailments and deep linguistic understanding.

📄 PDF Abstract BibTeX arXiv:2306.12213

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

fail 설명 없음

Similar Papers 제목 키워드 기반

Autocorrelations Decay in Texts and Applicability Limits of Language Models

2023-05-11 · Nikolay Mikhaylovskiy, Ilya Churilov

We show that the laws of autocorrelations decay in texts are closely related to applicability limits of language models. Using distributional semantics we empirically demonstrate that autocorrelations of words in texts d…

Beyond Context Limits: Subconscious Threads for Long-Horizon Reasoning

2025-07-22 · Hongyin Luo, Nathaniel Morgan, Tina Li, Derek Zhao 외 arxiv

To break the context limits of large language models (LLMs) that bottleneck reasoning accuracy and efficiency, we propose the Thread Inference Model (TIM), a family of LLMs trained for recursive and decompositional probl…

Information Retrieval

``ye word kis lang ka hai bhai?'' Testing the Limits of Word level Language Identification

2014-12-01 · WS 2014 12 · Sp Gella, ana, Kalika Bali, Monojit Choudhury
Language IdentificationTransliteration

A baseline revisited: Pushing the limits of multi-segment models for context-aware translation

2022-10-19 · Suvodeep Majumder, Stanislas Lauly, Maria Nadejde, Marcello Federico 외

This paper addresses the task of contextual translation using multi-segment models. Specifically we show that increasing model capacity further pushes the limits of this approach and that deeper models are more suited to…

Knowledge DistillationTranslation

Exploring the Limits of ChatGPT in Software Security Applications

2023-12-08 · Fangzhou Wu, Qingzhao Zhang, Ati Priya Bajaj, Tiffany Bao 외

Large language models (LLMs) have undergone rapid evolution and achieved remarkable results in recent times. OpenAI's ChatGPT, backed by GPT-3.5 or GPT-4, has gained instant popularity due to its strong capability across…

Vulnerability Detection