paper-with-me

Papers

Condition Number Analysis of Logistic Regression, and its Implications for Standard First-Order Solution Methods

2018-10-20 · Robert M. Freund, Paul Grigas, Rahul Mazumder

Logistic regression is one of the most popular methods in binary classification, wherein estimation of model parameters is carried out by solving the maximum likelihood (ML) optimization problem, and the ML estimator is defined to be the optimal solution of this problem. It is well known that the ML estimator exists when the data is non-separable, but fails to exist when the data is separable. First-order methods are the algorithms of choice for solving large-scale instances of the logistic regression problem. In this paper, we introduce a pair of condition numbers that measure the degree of non-separability or separability of a given dataset in the setting of binary classification, and we study how these condition numbers relate to and inform the properties and the convergence guarantees of first-order methods. When the training data is non-separable, we show that the degree of non-separability naturally enters the analysis and informs the properties and convergence guarantees of two standard first-order methods: steepest descent (for any given norm) and stochastic gradient descent. Expanding on the work of Bach, we also show how the degree of non-separability enters into the analysis of linear convergence of steepest descent (without needing strong convexity), as well as the adaptive convergence of stochastic gradient descent. When the training data is separable, first-order methods rather curiously have good empirical success, which is not well understood in theory. In the case of separable data, we demonstrate how the degree of separability enters into the analysis of $\ell_2$ steepest descent and stochastic gradient descent for delivering approximate-maximum-margin solutions with associated computational guarantees as well. This suggests that first-order methods can lead to statistically meaningful solutions in the separable case, even though the ML solution does not exist.

📄 PDF Abstract BibTeX arXiv:1810.08727

Code (0)

등록된 구현이 없습니다.

Tasks

Binary ClassificationGeneral Classificationregression

Methods 이 논문이 사용한 방법론

Logistic Regression Logistic Regression, despite its name, is a linear model for classification rather than regression. Logistic regression is also known in the literature as logit regression,…

Similar Papers 제목 키워드 기반

A Provably Accurate Randomized Sampling Algorithm for Logistic Regression

2024-02-26 · Agniva Chowdhury, Pradeep Ramuhalli

In statistics and machine learning, logistic regression is a widely-used supervised learning technique primarily employed for binary classification tasks. When the number of observations greatly exceeds the number of pre…

Binary Classificationregression

A Conditional Randomization Test for Sparse Logistic Regression in High-Dimension

2022-05-29 · Binh T. Nguyen, Bertrand Thirion, Sylvain Arlot

Identifying the relevant variables for a classification model with correct confidence levels is a central but difficult task in high-dimension. Despite the core role of sparse logistic regression in statistics and machin…

regressionVocal Bursts Intensity Prediction

Dimension-free uniform concentration bound for logistic regression

2024-05-28 · Shogo Nakakita

We provide a novel dimension-free uniform concentration bound for the empirical risk function of constrained logistic regression. Our bound yields a milder sufficient condition for a uniform law of large numbers than con…

regression

Sparse Logistic Regression Learns All Discrete Pairwise Graphical Models

2018-10-28 · NeurIPS 2019 12 · Shanshan Wu, Sujay Sanghavi, Alexandros G. Dimakis

We characterize the effectiveness of a classical algorithm for recovering the Markov graph of a general discrete pairwise graphical model from i.i.d. samples. The algorithm is (appropriately regularized) maximum conditio…

Allregression

An efficient model-free estimation of multiclass conditional probability

2012-09-22 · Tu Xu, Junhui Wang

Conventional multiclass conditional probability estimation methods, such as Fisher's discriminate analysis and logistic regression, often require restrictive distributional model assumption. In this paper, a model-free e…

quantile regressionregression