paper-with-me

홈 › Papers

Beyond the Commitment Boundary: Probing Epiphenomenal Chain-of-Thought in Large Reasoning Models

2026-06-11 · Daniel Scalena, Sara Candussio, Luca Bortolussi, Elisabetta Fersini, Malvina Nissim, Gabriele Sarti arxiv

Chain-of-thought (CoT) reasoning is the dominant paradigm for inference-time scaling in language models, yet the causal influence of individual steps on the final answer poorly understood. We estimate each step's causal importance via early exit and use this measure to study how answers form across the reasoning traces of several model families. Across diverse tasks, we find that reasoning typically crosses a \emph{commitment boundary} -- a sharp transition from transient intermediate guesses to a stable, high-confidence answer. This transition often happens in a single step, well before the model's reasoning block ends, and is followed by \emph{epiphenomenal} CoT steps that leave the final answer probability unaltered. Using attention probes, we show that answer-formation stages can be linearly decoded from intermediate reasoning steps with high accuracy and generalize robustly to unseen reasoning tasks. We exploit this signal to early-exit reasoning blocks at the commitment boundary, reducing the length of CoTs up to 55\% on average with negligible impact on model performance.

📄 PDF Abstract BibTeX arXiv:2606.13603

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Dynamics Within Latent Chain-of-Thought: An Empirical Study of Causal Structure

2026-02-09 · Zirui Li, Xuefeng Bai, Kehai Chen, Yizhi Li 외 arxiv

Latent or continuous chain-of-thought methods replace explicit textual rationales with a number of internal latent steps, but these intermediate computations are difficult to evaluate beyond correlation-based probes. In …

Beyond Prediction -- Structuring Epistemic Integrity in Artificial Reasoning Systems

2025-06-19 · Craig Steven Wright

This paper develops a comprehensive framework for artificial intelligence systems that operate under strict epistemic constraints, moving beyond stochastic language prediction to support structured reasoning, proposition…

Knowledge Graphs

Probing the Trajectories of Reasoning Traces in Large Language Models

2026-01-30 · Marthe Ballon, Brecht Verbeken, Vincent Ginis, Andres Algaba arxiv

Large language models (LLMs) increasingly solve difficult problems by producing "reasoning traces" before emitting a final response. However, it remains unclear how accuracy and decision commitment evolve along a reasoni…

VLADriveBench: Evaluating CoT-Action Relationship in VLA for Autonomous Driving

2026-06-10 · Thach Nguyen, Danhua Guo, Tom Lampo, Fei Wu 외 arxiv

Vision-language-action (VLA) models generate chain-of-thought (CoT) reasoning alongside driving trajectories, but existing benchmarks evaluate only trajectory quality and do not assess whether the CoT is relevant, consis…

Autonomous Driving

Path Drift in Large Reasoning Models:How First-Person Commitments Override Safety

2025-10-11 · Yuyi Huang, Runzhe Zhan, Lidia S. Chao, Ailin Tao 외 arxiv

As large language models (LLMs) are increasingly deployed for complex reasoning tasks, Long Chain-of-Thought (Long-CoT) prompting has emerged as a key paradigm for structured inference. Despite early-stage safeguards ena…