Estimating Population-Risk Curves Along Nonconvex Gradient Flows from the Training Sample
We estimate the conditional population-risk curve of a realized smooth nonconvex gradient flow from the training sample. Flow approximate leave-one-out (Flow-ALO) propagates a deletion response and evaluates omitted observations at approximate deleted paths. The risk-curve error decomposes into response approximation, exact-LOO fluctuation, and deletion-to-full risk transfer. On each fixed finite horizon, bounded centered training-loss gradients, a one-sided Hessian lower bound, locally Lipschitz Hessians, and a strict tube-closure condition yield an explicit $(n-1)^{-2}$ bound for the deletion-response error. Bounded evaluation-loss gradients transfer the deletion-response bound to the score without requiring the Hessian to be invertible. Direct first-order jackknife cancellation and exact-LOO concentration control deletion-to-full risk transfer and fluctuation, respectively, completing recovery of the conditional population-risk curve. For bounded smooth two-layer mean-field networks training both layers, the score-error bound is uniform in width.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
A Bayesian Nonparametric Approach for Estimating Individualized Treatment-Response Curves
We study the problem of estimating the continuous response over time to interventions using observational time series---a retrospective dataset where the policy by which the data are generated is unknown to the learner. …
Decision MakingKidney FunctionTime SeriesTime Series AnalysisThe Health Status of a Population: Health State and Survival Curves, and HALE Estimates
In this paper we explore the very important case of finding a health measure in the lines of the survival curve but independent of the standard deviation parameter. This is done by estimating the health state curve and c…
On the Local Minima of the Empirical Risk
Population risk is always of primary interest in machine learning; however, learning algorithms only have access to the empirical risk. Even for applications with nonconvex nonsmooth losses (such as modern deep networks)…
Explainable Machine Learning for Pediatric Dental Risk Stratification Using Socio-Demographic Determinants
Background: Pediatric dental disease remains one of the most prevalent and inequitable chronic health conditions worldwide. Although strong epidemiological evidence links oral health outcomes to socio-economic and demogr…
Differentially Private SGDA for Minimax Problems
Stochastic gradient descent ascent (SGDA) and its variants have been the workhorse for solving minimax problems. However, in contrast to the well-studied stochastic gradient descent (SGD) with differential privacy (DP) c…