paper-with-me

Papers

Probing the Trajectories of Reasoning Traces in Large Language Models

2026-01-30 · Marthe Ballon, Brecht Verbeken, Vincent Ginis, Andres Algaba arxiv

Large language models (LLMs) increasingly solve difficult problems by producing "reasoning traces" before emitting a final response. However, it remains unclear how accuracy and decision commitment evolve along a reasoning trajectory, and whether intermediate trace segments provide answer-relevant information beyond generic length or stylistic effects. Here, we propose a protocol to systematically probe the trajectories of reasoning traces in LLMs by 1) generating a model's reasoning trace, 2) truncating it at fixed token-percentiles, and 3) injecting each partial trace back into the model (or a different model) to measure the induced distribution over answer choices via next-token probabilities. We apply this protocol to the open-source Qwen3-4B/-8B/-14B and gpt-oss-20b/-120b models across the multiple-choice GPQA Diamond and MMLU-Pro benchmarks. We find that accuracy and decision commitment consistently increase as the percentage of provided reasoning tokens grows. These gains are primarily driven by relevant content in the model generation rather than context length or generic "reasoning style" effects. Stronger models often backtrack successfully from incorrect partial traces, but immediate answers often remain anchored in the weaker model's incorrect response. More broadly, we show that trajectory probing provides diagnostics for efficient and safer deployment of reasoning models as the measurements can inform practical trace-handling and monitoring policies that improve reliability without assuming intermediate tokens are inherently faithful explanations.

📄 PDF Abstract BibTeX arXiv:2601.23163

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Thought-Level Beam Search for Reasoning

2026-08-11 · Lijie Yang, Hongyin Luo, Jiawei Zhao, Tri Dao 외 hf

Test-time compute scaling is a primary driver of performance in large reasoning models (LRMs), but extreme inefficiency bounds current approaches, shifting the critical question from how much compute to spend, to where t…

Leaky Thoughts: Large Reasoning Models Are Not Private Thinkers

2025-06-18 · Tommaso Green, Martin Gubri, Haritz Puerto, Sangdoo Yun 외

We study privacy leakage in the reasoning traces of large reasoning models used as personal agents. Unlike final outputs, reasoning traces are often assumed to be internal and safe. We challenge this assumption by showin…

Revisiting Complete Reasoning Traces for Post-Training

2026-09-07 · Jaehui Hwang, Sangdoo Yun, Byeongho Heo, Dongyoon Han hf

Large language models (LLMs) are often post-trained on pre-collected reasoning trajectories to improve their reasoning capability. Such trajectories tend to be long due to complex, interwoven paths, which often include d…

Reinforcement Learning

Confidence Geometry Reveals Trace-Level Correctness in Large Language Model Reasoning

2026-05-16 · Shuo Liu, Ding Liu, Shi-Ju Ran arxiv

Large language models (LLMs) generate not only reasoning text, but also token-level confidence trajectories that record how uncertainty evolves during inference. Whether these trajectories are relevant to reasoning corre…

InfoDensity: Rewarding Information-Dense Traces for Efficient Reasoning

2026-03-18 · Chengwei Wei, Jung-jae Kim, Longyin Zhang, Shengkai Chen 외 arxiv

Large Language Models (LLMs) with extended reasoning capabilities often generate verbose and redundant reasoning traces, incurring unnecessary computational cost. While existing reinforcement learning approaches address …

Reinforcement Learning