paper-with-me

홈 › Papers

Delving into the Reversal Curse: How Far Can Large Language Models Generalize?

2024-10-24 · Zhengkai Lin, Zhihang Fu, Kai Liu, Liang Xie, Binbin Lin, Wenxiao Wang, Deng Cai, Yue Wu, Jieping Ye

While large language models (LLMs) showcase unprecedented capabilities, they also exhibit certain inherent limitations when facing seemingly trivial tasks. A prime example is the recently debated "reversal curse", which surfaces when models, having been trained on the fact "A is B", struggle to generalize this knowledge to infer that "B is A". In this paper, we examine the manifestation of the reversal curse across various tasks and delve into both the generalization abilities and the problem-solving mechanisms of LLMs. This investigation leads to a series of significant insights: (1) LLMs are able to generalize to "B is A" when both A and B are presented in the context as in the case of a multiple-choice question. (2) This generalization ability is highly correlated to the structure of the fact "A is B" in the training documents. For example, this generalization only applies to biographies structured in "[Name] is [Description]" but not to "[Description] is [Name]". (3) We propose and verify the hypothesis that LLMs possess an inherent bias in fact recalling during knowledge application, which explains and underscores the importance of the document structure to successful learning. (4) The negative impact of this bias on the downstream performance of LLMs can hardly be mitigated through training alone. These findings offer a novel perspective on interpreting LLMs' generalization through their intrinsic mechanisms and provide insights for developing more effective learning methods. Our code and data are available at https://github.com/alibaba/thinking_bias.git.

📄 PDF Abstract BibTeX arXiv:2410.18808

Code (1)

alibaba/thinking_bias 공식 구현 pytorch

Tasks

Multiple-choice

Similar Papers 제목 키워드 기반

An Analysis and Mitigation of the Reversal Curse

2023-11-13 · Ang Lv, Kaiyi Zhang, Shufang Xie, Quan Tu 외

Recent research observed a noteworthy phenomenon in large language models (LLMs), referred to as the ``reversal curse.'' The reversal curse is that when dealing with two entities, denoted as $a$ and $b$, connected by the…

DenoisingLanguage Modelling

Mitigating Reversal Curse in Large Language Models via Semantic-aware Permutation Training

2024-03-01 · Qingyan Guo, Rui Wang, Junliang Guo, Xu Tan 외

While large language models (LLMs) have achieved impressive performance across diverse tasks, recent studies showcase that causal LLMs suffer from the "reversal curse". It is a typical example that the model knows "A's f…

Language Modelling

The Factorization Curse: Which Tokens You Predict Underlie the Reversal Curse and More

2024-06-07 · Ouail Kitouni, Niklas Nolte, Diane Bouchacourt, Adina Williams 외

Today's best language models still struggle with hallucinations: factually incorrect generations, which impede their ability to reliably retrieve information seen during training. The reversal curse, where models cannot …

Information RetrievalRetrieval

DiffER: Diffusion Entity-Relation Modeling for Reversal Curse in Diffusion Large Language Models

2026-01-12 · Shaokai He, Kaiwen Wei, Xinyi Zeng, Xiang Chen 외 arxiv

The "reversal curse" refers to the phenomenon where large language models (LLMs) exhibit predominantly unidirectional behavior when processing logically bidirectional relationships. Prior work attributed this to autoregr…

Breaking the Reversal Curse in Autoregressive Language Models via Identity Bridge

2026-02-02 · Xutao Ma, Yixiao Huang, Hanlin Zhu, Somayeh Sojoudi arxiv

Autoregressive large language models (LLMs) have achieved remarkable success in many complex tasks, yet they can still fail in very simple logical reasoning such as the "reversal curse" -- when trained on forward knowled…

Logical Reasoning