paper-with-me

홈 › Papers

TPV: Parameter Perturbations Through the Lens of Test Prediction Variance

2025-12-11 · Devansh Arpit arxiv

We introduce test prediction variance (TPV)--the first-order sensitivity of a trained model's outputs to parameter perturbations--as a unifying framework for analyzing post-training robustness. TPV is a fully label-free object whose trace form separates the geometry of the trained model from the specific perturbation mechanism, placing SGD noise, label noise, quantization, and pruning under a single lens. The resulting expressions recover the wide-minima hypothesis for SGD and quantization noise, and yield a distinct Jacobian-spectral characterization for label noise connecting label-noise TPV with benign overfitting in nonlinear networks. Theoretically, we prove that training-set TPV converges to its test-set counterpart in the overparameterized limit, irrespective of generalization performance, providing the first result that prediction variance under local parameter perturbations can be inferred from training inputs alone. Empirically, this stability holds far more broadly, including at very low widths. Further, TPV correlates well with test loss, enabling practical applications: JBR, a label-free pruning criterion derived from TPV geometry matching state-of-the-art baselines; and training-set based model selection signal for in-distribution and transfer learning scenarios. Code available at github.com/devansharpit/TPV.

📄 PDF Abstract BibTeX arXiv:2512.11089

Code (0)

등록된 구현이 없습니다.

Tasks

Transfer Learning

Similar Papers 제목 키워드 기반

Adaptive Camera Sensor for Vision Models

2025-03-04 · Eunsu Baek, Sunghwan Han, Taesik Gong, Hyung-Sin Kim

Domain shift remains a persistent challenge in deep-learning-based computer vision, often requiring extensive model modifications or large labeled datasets to address. Inspired by human visual perception, which adjusts i…

Adversarial Robustness through the Lens of Convolutional Filters

2022-04-05 · Paul Gavrikov, Janis Keuper

Deep learning models are intrinsically sensitive to distribution shifts in the input data. In particular, small, barely perceivable perturbations to the input data can force models to make wrong predictions with high con…

Adversarial Robustness

Bayesian Strong Gravitational-Lens Modeling on Adaptive Grids: Objective Detection of Mass Substructure in Galaxies

2008-05-02 · S. Vegetti, L. V. E. Koopmans

We introduce a new adaptive and fully Bayesian grid-based method to model strong gravitational lenses with extended images. The primary goal of this method is to quantify the level of luminous and dark-mass substructure …

Non-robust Features through the Lens of Universal Perturbations

2021-01-01 · Sung Min Park, Kuo-An Wei, Kai Yuanqing Xiao, Jerry Li 외

Recent work ties adversarial perturbations to so-called non-robust features. These are features which are susceptible to small perturbations and believed to be incomprehensible to humans, but still useful for (generaliza…

Eliciting Latent Predictions from Transformers with the Tuned Lens

2023-03-14 · Nora Belrose, Zach Furman, Logan Smith, Danny Halawi 외

We analyze transformers from the perspective of iterative inference, seeking to understand how model predictions are refined layer by layer. To do so, we train an affine probe for each block in a frozen pretrained model,…

Language Modelling