paper-with-me

Papers

The Last Word Often Wins: A Format Confound in Chain-of-Thought Corruption Studies

2026-05-11 · Gabriel Garcia arxiv

Corruption studies, the standard tool for evaluating chain-of-thought (CoT) faithfulness, infer which steps are ``computationally important'' from accuracy loss when steps are corrupted. We show that when benchmark chains end with an explicit terminal answer line, as in GSM8K and MATH, these tests largely measure \emph{answer placement} rather than where intermediate computation is carried out. Using matched GSM8K examples, removing only the final answer statement while preserving all reasoning collapses suffix sensitivity by about $19\times$ for Qwen~2.5-3B ($N{=}300$, $p{=}0.022$). Conflicting-answer prompts, which contain correct reasoning but a wrong explicit final answer, drive accuracy to zero or near-zero at 7B across five open-weight model families; wrong-answer following is strong at 3B--7B and attenuates sharply at larger scales. Replications on MATH, within-stable comparisons at 7B, and suffix-free chains show the same pattern in different guises: corruption sensitivity tracks the location of explicit answer text, not a fixed computational depth in the reasoning. Generation-time probes indicate that final answers are rarely early-determined during generation (${<}5\%$ early commitment), yet consumption-time behavior systematically follows explicit answer text. The confound is therefore largely a readout effect when the chain is consumed. We propose a three-prerequisite protocol (question-only control, format characterization, and an all-position sweep) as a practical minimum for future corruption-based faithfulness studies.

📄 PDF Abstract BibTeX arXiv:2605.10799

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Causal Falsification of Digital Twins

2023-01-17 · Rob Cornish, Muhammad Faaiz Taufiq, Arnaud Doucet, Chris Holmes

Digital twins are virtual systems designed to predict how a real-world process will evolve in response to interventions. This modelling paradigm holds substantial promise in many applications, but rigorous procedures for…

Causal Inference

Disambiguatory Signals are Stronger in Word-initial Positions

2021-02-03 · EACL 2021 2 · Tiago Pimentel, Ryan Cotterell, Brian Roark

Psycholinguistic studies of human word processing and lexical access provide ample evidence of the preferred nature of word-initial versus word-final segments, e.g., in terms of attention paid by listeners (greater) or t…

Informativeness

Detecting Unobserved Confounders: A Kernelized Regression Approach

2026-01-01 · Yikai Chen, Yunxin Mao, Chunyuan Zheng, Hao Zou 외 arxiv

Detecting unobserved confounders is crucial for reliable causal inference in observational studies. Existing methods require either linearity assumptions or multiple heterogeneous environments, limiting applicability to …

Computational EfficiencyCausal Inference

Leading Whitespaces of Language Models' Subword Vocabulary Pose a Confound for Calculating Word Probabilities

2024-06-16 · Byung-Doh Oh, William Schuler

Predictions of word-by-word conditional probabilities from Transformer-based language models are often evaluated to model the incremental processing difficulty of human readers. In this paper, we argue that there is a co…

An Overview of Digital Twins Application Domains in Smart Energy Grid

2021-04-16 · Tudor Cioara, Ionut Anghel, Marcel Antal, Ioan Salomie 외

The Digital Twins offer promising solutions for smart grid challenges related to the optimal operation, management, and control of energy assets, for safe and reliable distribution of energy. These challenges are more pr…

Cloud Computingenergy managementManagement