paper-with-me

홈 › Papers

Asymptotic normality of robust risk minimizers

2020-04-05 · Stanislav Minsker

This paper investigates asymptotic properties of algorithms that can be viewed as robust analogues of the classical empirical risk minimization. These strategies are based on replacing the usual empirical average by a robust proxy of the mean, such as the (version of) the median of means estimator. It is well known by now that the excess risk of resulting estimators often converges to zero at optimal rates under much weaker assumptions than those required by their ``classical'' counterparts. However, less is known about the asymptotic properties of the estimators themselves, for instance, whether robust analogues of the maximum likelihood estimators are asymptotically efficient. We make a step towards answering these questions and show that for a wide class of parametric problems, minimizers of the appropriately defined robust proxy of the risk converge to the minimizers of the true risk at the same rate, and often have the same asymptotic variance, as the estimators obtained by minimizing the usual empirical risk.

📄 PDF Abstract BibTeX arXiv:2004.02328

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Asyptotic Normality for Maximum Likelihood Estimation and Operational Risk

2016-08-25

Operational risk models commonly employ maximum likelihood estimation (MLE) to fit loss data to heavy-tailed distributions. Yet several desirable properties of MLE (e.g. asymptotic normality) are generally valid only for…

valid

Non-convex learning via Stochastic Gradient Langevin Dynamics: a nonasymptotic analysis

2017-02-13 · Maxim Raginsky, Alexander Rakhlin, Matus Telgarsky

Stochastic Gradient Langevin Dynamics (SGLD) is a popular variant of Stochastic Gradient Descent, where properly scaled isotropic Gaussian noise is added to an unbiased estimate of the gradient at each iteration. This mo…

Post Reinforcement Learning Inference

2023-02-17 · Vasilis Syrgkanis, Ruohan Zhan

We consider estimation and inference using data collected from reinforcement learning algorithms. These algorithms, characterized by their adaptive experimentation, interact with individual units over multiple stages, dy…

counterfactualOff-policy evaluationreinforcement-learningReinforcement Learning+1

On Binary Classification in Extreme Regions

2018-12-01 · NeurIPS 2018 12 · Hamid Jalalzai, Stephan Clémençon, Anne Sabourin

In pattern recognition, a random label Y is to be predicted based upon observing a random vector X valued in $\mathbb{R}^d$ with d>1 by means of a classification rule with minimum probability of error. In a wide variety …

Binary ClassificationClassificationGeneral Classification

On the Efficiency of ERM in Feature Learning

2024-11-18 · Ayoub El Hanchi, Chris J. Maddison, Murat A. Erdogdu

Given a collection of feature maps indexed by a set $\mathcal{T}$, we study the performance of empirical risk minimization (ERM) on regression problems with square loss over the union of the linear classes induced by the…