paper-with-me

홈 › Papers

When Is Thinking Enough? Early Exit via Sufficiency Assessment for Efficient Reasoning

2026-04-08 · Yang Xiang, Yixin Ji, Ruotao Xu, Dan Qiao, Zheming Yang, Juntao Li, Min Zhang arxiv

Large reasoning models (LRMs) have achieved remarkable performance in complex reasoning tasks, driven by their powerful inference-time scaling capability. However, LRMs often suffer from overthinking, which results in substantial computational redundancy and significantly reduces efficiency. Early-exit methods aim to mitigate this issue by terminating reasoning once sufficient evidence has been generated, yet existing approaches mostly rely on handcrafted or empirical indicators that are unreliable and impractical. In this work, we introduce Dynamic Thought Sufficiency in Reasoning (DTSR), a novel framework for efficient reasoning that enables the model to dynamically assess the sufficiency of its chain-of-thought (CoT) and determine the optimal point for early exit. Inspired by human metacognition, DTSR operates in two stages: (1) Reflection Signal Monitoring, which identifies reflection signals as potential cues for early exit, and (2) Thought Sufficiency Check, which evaluates whether the current CoT is sufficient to derive the final answer. Experimental results on the Qwen3 models show that DTSR reduces reasoning length by 28.9%-34.9% with minimal performance loss, effectively mitigating overthinking. We further discuss overconfidence in LRMs and self-evaluation paradigms, providing valuable insights for early-exit reasoning.

📄 PDF Abstract BibTeX arXiv:2604.06787

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Rethinking Calibration for Early-Exit Neural Networks

2025-08-29 · Piotr Kubaty, Filip Szatkowski, Grzegorz Choczyński, Eric Nalisnick 외 arxiv

Early-exit neural networks (EENNs) accelerate inference by allowing intermediate classifiers to stop computation once predictions are confident enough. Most methods rely on confidence thresholds for exiting, and conseque…

Classifier calibration

How Early Is Early Enough? Design-Dependent Observation-Window Sufficiency in Subscription Churn Prediction

2026-07-01 · Xiao Han, Yao Xiao, Chenyu Wu, Tongchen Zhang arxiv

How many days of early behavior suffice for subscription churn prediction? In the public KKBox dataset, the early indicator of churn is typically an indicator of someone's contract status; however, when looking in the he…

S-GRPO: Early Exit via Reinforcement Learning in Reasoning Models

2025-05-12 · Muzhi Dai, Chenxu Yang, Qingyi Si

As Test-Time Scaling emerges as an active research focus in the large language model community, advanced post-training methods increasingly emphasize extending chain-of-thought (CoT) generation length, thereby enhancing …

GSM8KLarge Language ModelMathreinforcement-learning+1

SuCo: Sufficiency-guided Continuous Adaptive Reasoning

2026-06-16 · Jiahao Wang, Bingyu Liang, Chenhao Hu, Longhui Zhang 외 arxiv

Despite remarkable performance on complex tasks, Large Reasoning Models (LRMs) often generate excessively long Chain-of-Thoughts (CoT), inflating computational costs even for simple queries. Existing efforts to mitigate …

Reinforcement Learning

The Zero-Step Thinking: An Empirical Study of Mode Selection as Harder Early Exit in Reasoning Models

2025-10-22 · Yuqiao Tan, Shizhu He, Kang Liu, Jun Zhao arxiv

Reasoning models have demonstrated exceptional performance in tasks such as mathematics and logical reasoning, primarily due to their ability to engage in step-by-step thinking during the reasoning process. However, this…

Logical Reasoning