paper-with-me

Papers

Optimal lower bounds for logistic log-likelihoods

2024-10-14 · Niccolò Anceschi, Tommaso Rigon, Giacomo Zanella, Daniele Durante

The logit transform is arguably the most widely-employed link function beyond linear settings. This transformation routinely appears in regression models for binary data and provides, either explicitly or implicitly, a core building-block within state-of-the-art methodologies for both classification and regression. Its widespread use, combined with the lack of analytical solutions for the optimization of general losses involving the logit transform, still motivates active research in computational statistics. Among the directions explored, a central one has focused on the design of tangent lower bounds for logistic log-likelihoods that can be tractably optimized, while providing a tight approximation of these log-likelihoods. Although progress along these lines has led to the development of effective minorize-maximize (MM) algorithms for point estimation and coordinate ascent variational inference schemes for approximate Bayesian inference under several logit models, the overarching focus in the literature has been on tangent quadratic minorizers. In fact, it is still unclear whether tangent lower bounds sharper than quadratic ones can be derived without undermining the tractability of the resulting minorizer. This article addresses such a challenging question through the design and study of a novel piece-wise quadratic lower bound that uniformly improves any tangent quadratic minorizer, including the sharpest ones, while admitting a direct interpretation in terms of the classical generalized lasso problem. As illustrated in a ridge logistic regression, this unique connection facilitates more effective implementations than those provided by available piece-wise bounds, while improving the convergence speed of quadratic ones.

📄 PDF Abstract BibTeX arXiv:2410.10309

Code (0)

등록된 구현이 없습니다.

Tasks

Bayesian InferenceregressionVariational Inference

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…
Focus 설명 없음
Variational Inference 설명 없음

Similar Papers 제목 키워드 기반

Quadruply Stochastic Gaussian Processes

2020-06-04 · Trefor W. Evans, Prasanth B. Nair

We introduce a stochastic variational inference procedure for training scalable Gaussian process (GP) models whose per-iteration complexity is independent of both the number of training points, $n$, and the number basis …

Gaussian ProcessesregressionStochastic OptimizationVariational Inference

Minimax optimality of deep neural networks on dependent data via PAC-Bayes bounds

2024-10-29 · Pierre Alquier, William Kengne

In a groundbreaking work, Schmidt-Hieber (2020) proved the minimax optimality of deep neural networks with ReLu activation for least-square regression estimation over a large class of functions defined by composition. In…

regression

Belief likelihood function for generalised logistic regression

2018-08-07 · Fabio Cuzzolin

The notion of belief likelihood function of repeated trials is introduced, whenever the uncertainty for individual trials is encoded by a belief measure (a finite random set). This generalises the traditional likelihood …

regression

The Space Complexity of Approximating Logistic Loss

2024-12-03 · Gregory Dexter, Petros Drineas, Rajiv Khanna

We provide space complexity lower bounds for data structures that approximate logistic loss up to $\epsilon$-relative error on a logistic regression problem with data $\mathbf{X} \in \mathbb{R}^{n \times d}$ and labels $…

Optimal Dimension-Free Sampling for Regularized Classification

2026-05-22 · Meysam Alishahi, Alexander Munteanu, Simon Omlor, Jeff M. Phillips arxiv

We prove optimal sampling bounds achieving $(1\pm\varepsilon)$-relative error for a broad class of Lipschitz continuous classification loss functions under various regularization terms. This includes important functions …