paper-with-me

홈 › Papers

Wild refitting for black box prediction

2025-06-26 · Martin J. Wainwright

We describe and analyze a computionally efficient refitting procedure for computing high-probability upper bounds on the instance-wise mean-squared prediction error of penalized nonparametric estimates based on least-squares minimization. Requiring only a single dataset and black box access to the prediction method, it consists of three steps: computing suitable residuals, symmetrizing and scaling them with a pre-factor $\rho$, and using them to define and solve a modified prediction problem recentered at the current estimate. We refer to it as wild refitting, since it uses Rademacher residual symmetrization as in a wild bootstrap variant. Under relatively mild conditions allowing for noise heterogeneity, we establish a high probability guarantee on its performance, showing that the wild refit with a suitably chosen wild noise scale $\rho$ gives an upper bound on prediction error. This theoretical analysis provides guidance into the design of such procedures, including how the residuals should be formed, the amount of noise rescaling in the wild sub-problem needed for upper bounds, and the local stability properties of the block-box procedure. We illustrate the applicability of this procedure to various problems, including non-rigid structure-from-motion recovery with structured matrix penalties; plug-and-play image restoration with deep neural network priors; and randomized sketching with kernel methods.

📄 PDF Abstract BibTeX arXiv:2506.21460

Code (0)

등록된 구현이 없습니다.

Tasks

Image RestorationPrediction

Similar Papers 제목 키워드 기반

Interleaved Resampling and Refitting: Data and Compute-Efficient Evaluation of Black-Box Predictors

2026-03-15 · Haichen Hu, David Simchi-Levi arxiv

We study the problem of evaluating the excess risk of large-scale empirical risk minimization under the square loss. Leveraging the idea of wild refitting and resampling, we assume only black-box access to the training a…

Perturbing the Derivative: Doubly Wild Refitting for Model-Free Evaluation of Opaque Machine Learning Predictors

2025-11-24 · Haichen Hu, David Simchi-Levi arxiv

We study the problem of excess risk evaluation for empirical risk minimization (ERM) under convex losses. We show that by leveraging the idea of wild refitting, one can upper bound the excess risk through the so-called "…

Perturbing the Derivative: Wild Refitting for Model-Free Evaluation of Machine Learning Models under Bregman Losses

2025-09-02 · Haichen Hu, David Simchi-Levi arxiv

We study the excess risk evaluation of classical penalized empirical risk minimization (ERM) with Bregman losses. We show that by leveraging the idea of wild refitting, one can efficiently upper bound the excess risk thr…

The Lifecycle of a Statistical Model: Model Failure Detection, Identification, and Refitting

2022-02-08 · Alnur Ali, Maxime Cauchois, John C. Duchi

The statistical machine learning community has demonstrated considerable resourcefulness over the years in developing highly expressive tools for estimation, prediction, and inference. The bedrock assumptions underlying …

modelTime SeriesTime Series Analysis

EVARS-GPR: EVent-triggered Augmented Refitting of Gaussian Process Regression for Seasonal Data

2021-07-06 · Florian Haselbeck, Dominik G. Grimm

Time series forecasting is a growing domain with diverse applications. However, changes of the system behavior over time due to internal or external influences are challenging. Therefore, predictions of a previously lear…

Change Point DetectionData AugmentationGPRregression+3