paper-with-me

홈 › Papers

Matrix-Decoupled Concentration for Autoregressive Sequences: Dimension-Free Guarantees for Sparse Long-Context Rewards

2026-05-07 · Pei-Sen Li arxiv

Sequence-level evaluations in autoregressive Large Language Models (LLMs) rely on highly dependent token generation. Establishing tight concentration bounds for these processes remains a challenge due to two fundamental bottlenecks in existing frameworks: (i) classical inequalities typically separate dependency structures from target sensitivities, leading to a scalar collapse that inflates the variance proxy to a suboptimal $\mathcal{O}(N)$ for sparse terminal rewards; (ii) conversely, while certain spatial methods achieve tighter bounds, they lack the strictly causal filtration required by sequential generation, rendering them inapplicable to the autoregressive setting. To resolve both bottlenecks, we establish a sharp McDiarmid-type inequality for dependent sequences, governed strictly by the exact matrix-vector multiplication of the causal dependency resolvent and the target sensitivity vector. This Matrix-Decoupled Concentration (MDC) framework natively recovers optimal constants for Markov chains and exploits directed $d$-separation to yield order-optimal bounds for causal trees. Crucially, by exactly preserving the coordinate-wise sparsity of rewards within a strictly causal framework, MDC mathematically prevents scalar collapse, guaranteeing a dimension-free $\mathcal{O}(1)$ variance proxy and providing a rigorous mathematical justification for the stability of long-context reasoning.

📄 PDF Abstract BibTeX arXiv:2605.06017

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Missing Data in Sparse Transition Matrix Estimation for Sub-Gaussian Vector Autoregressive Processes

2018-02-26 · Amin Jalali, Rebecca Willett

High-dimensional time series data exist in numerous areas such as finance, genomics, healthcare, and neuroscience. An unavoidable aspect of all such datasets is missing data, and dealing with this issue has been an impor…

Time SeriesTime Series Analysis

A Finite-Sample Deviation Bound for Stable Autoregressive Processes

2019-12-17 · L4DC 2020 6 · Rodrigo A. González, Cristian R. Rojas

In this paper, we study non-asymptotic deviation bounds of the least squares estimator in Gaussian AR($n$) processes. By relying on martingale concentration inequalities and a tail-bound for $\chi^2$ distributed variable…

A Bayesian sparse factor model with adaptive posterior concentration

2023-05-29 · Ilsang Ohn, Lizhen Lin, Yongdai Kim

In this paper, we propose a new Bayesian inference method for a high-dimensional sparse factor model that allows both the factor dimensionality and the sparse structure of the loading matrix to be inferred. The novelty i…

Bayesian Inference

Dimension-free Concentration Bounds on Hankel Matrices for Spectral Learning

2013-12-21 · François Denis, Mattias Gybels, Amaury Habrard

Learning probabilistic models over strings is an important issue for many applications. Spectral methods propose elegant solutions to the problem of inferring weighted automata from finite samples of variable-length stri…

Concentration inequalities for high-dimensional linear processes with dependent innovations

2023-07-23 · Eduardo Fonseca Mendes, Fellipe Lopes

We develop concentration inequalities for the $l_\infty$ norm of vector linear processes with sub-Weibull, mixingale innovations. This inequality is used to obtain a concentration bound for the maximum entrywise norm of …

Time Series