paper-with-me

홈 › Papers

Harnessing the Unseen: The Hidden Influence of Intrinsic Knowledge in Long-Context Language Models

2025-04-11 · Yu Fu, HAZ Sameen Shahgir, Hui Liu, Xianfeng Tang, Qi He, Yue Dong

Recent advances in long-context models (LCMs), designed to handle extremely long input contexts, primarily focus on utilizing external contextual information, often leaving the influence of large language models' intrinsic knowledge underexplored. In this work, we investigate how this intrinsic knowledge affects content generation and demonstrate that its impact becomes increasingly pronounced as context length extends. Furthermore, we show that the model's ability to utilize intrinsic knowledge, which we call intrinsic retrieval ability, does not improve simultaneously with its ability to leverage contextual knowledge through extrinsic retrieval ability. Moreover, better extrinsic retrieval can interfere with the model's ability to use its own knowledge effectively, limiting its full potential. To bridge this gap, we design a simple yet effective Hybrid Needle-in-a-Haystack test that evaluates models based on their capabilities across both retrieval abilities, rather than solely emphasizing extrinsic retrieval ability. Our experimental results reveal that Qwen-2.5 models significantly outperform Llama-3.1 models, demonstrating superior intrinsic retrieval ability. Moreover, even the more powerful Llama-3.1-70B-Instruct model fails to exhibit better performance under LCM conditions, highlighting the importance of evaluating models from a dual-retrieval perspective.

📄 PDF Abstract BibTeX arXiv:2504.08202

Code (0)

등록된 구현이 없습니다.

Tasks

Retrieval

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

Harnessing the Intrinsic Knowledge of Pretrained Language Models for Challenging Text Classification Settings

2024-08-28 · Lingyu Gao

Text classification is crucial for applications such as sentiment analysis and toxic text filtering, but it still faces challenges due to the complexity and ambiguity of natural language. Recent advancements in deep lear…

In-Context LearningSentiment Analysistext-classificationText Classification

Learning from the Unseen: Offline Reinforcement Learning with Hidden Actions

2026-07-28 · Zeyu Bian, Ying Zhou, Yifan Cui arxiv

Standard offline reinforcement learning (RL) algorithms typically assume that the actions in the dataset are observed without error. However, in many real-world applications, the true actions are unobserved and only nois…

Reinforcement LearningOffline RL

Proactive Adversarial Defense: Harnessing Prompt Tuning in Vision-Language Models to Detect Unseen Backdoored Images

2024-12-11 · Kyle Stein, Andrew Arash Mahyari, Guillermo Francia, Eman El-Sheikh

Backdoor attacks pose a critical threat by embedding hidden triggers into inputs, causing models to misclassify them into target labels. While extensive research has focused on mitigating these attacks in object recognit…

Adversarial Defensebackdoor defenseObject Recognition

Joint graph learning from Gaussian observations in the presence of hidden nodes

2022-12-04 · Samuel Rey, Madeline Navarro, Andrei Buciulea, Santiago Segarra 외

Graph learning problems are typically approached by focusing on learning the topology of a single graph when signals from all nodes are available. However, many contemporary setups involve multiple related networks and, …

Graph LearningGraph Similarity

Intrinsic Self-correction for Enhanced Morality: An Analysis of Internal Mechanisms and the Superficial Hypothesis

2024-07-21 · Guangliang Liu, Haitao Mao, Jiliang Tang, Kristen Marie Johnson

Large Language Models (LLMs) are capable of producing content that perpetuates stereotypes, discrimination, and toxicity. The recently proposed moral self-correction is a computationally efficient method for reducing har…

Question AnsweringText Generation