paper-with-me

홈 › Papers

Internal Bias in Reasoning Models leads to Overthinking

2025-05-22 · Renfei Dang, ShuJian Huang, Jiajun Chen

While current reasoning models possess strong exploratory capabilities, they are often criticized for overthinking due to redundant and unnecessary reflections. In this work, we reveal for the first time that overthinking in reasoning models may stem from their internal bias towards input texts. Upon encountering a reasoning problem, the model immediately forms a preliminary guess about the answer, which we term as an internal bias since it is not derived through actual reasoning. When this guess conflicts with its reasoning result, the model tends to engage in reflection, leading to the waste of computational resources. Through further interpretability experiments, we find that this behavior is largely driven by the model's excessive attention to the input section, which amplifies the influence of internal bias on its decision-making process. Additionally, by masking out the original input section, the affect of internal bias can be effectively alleviated and the reasoning length could be reduced by 31%-53% across different complex reasoning tasks. Notably, in most cases, this approach also leads to improvements in accuracy. These findings demonstrate a causal relationship between internal bias and overthinking.

📄 PDF Abstract BibTeX arXiv:2505.16448

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음

Similar Papers 제목 키워드 기반

The Danger of Overthinking: Examining the Reasoning-Action Dilemma in Agentic Tasks

2025-02-12 · Alejandro Cuadron, Dacheng Li, Wenjie Ma, Xingyao Wang 외

Large Reasoning Models (LRMs) represent a breakthrough in AI problem-solving capabilities, but their effectiveness in interactive environments can be limited. This paper introduces and analyzes overthinking in LRMs. A ph…

Reconsidering Overthinking: Penalizing Internal and External Redundancy in CoT Reasoning

2025-08-04 · Taihang Zhen, Jialiang Hong, Kai Chen, Guang Yang 외 arxiv

Large reasoning models (LRMs) often exhibit overthinking, producing verbose Chain-of-Thought (CoT) traces that increase inference cost and obscure the underlying reasoning process. Existing CoT compression methods mainly…

Reinforcement LearningSemantic Similarity

Quantized Reasoning Models Think They Need to Think Longer, but They Do Not

2026-05-29 · Sanae Lotfi, Polina Kirichenko, Steven Li, Zechun Liu arxiv

Post-training quantization (PTQ) is widely used to deploy large language models efficiently, but its effect on reasoning models is not well understood. Across math, coding, and science QA, we find that aggressive PTQ red…

DRQA: Dynamic Reasoning Quota Allocation for Controlling Overthinking in Reasoning Large Language Models

2025-08-25 · Kaiwen Yan, Xuanqing Shi, Hongcheng Guo, Wenxuan Wang 외 arxiv

Reasoning large language models (RLLMs), such as OpenAI-O3 and DeepSeek-R1, have recently demonstrated remarkable capabilities by performing structured and multi-step reasoning. However, recent studies reveal that RLLMs …

Reinforcement Learning

The Evolution of Thought: Tracking LLM Overthinking via Reasoning Dynamics Analysis

2025-08-25 · Zihao Wei, Liang Pang, Jiahao Liu, Wenjie Shi 외 arxiv

Test-time scaling via explicit reasoning trajectories significantly boosts large language model (LLM) performance but often triggers overthinking. To explore this, we analyze reasoning through two lenses: Reasoning Lengt…