paper-with-me

홈 › Papers

Learning with Smooth Hinge Losses

2021-02-27 · JunRu Luo, Hong Qiao, Bo Zhang

Due to the non-smoothness of the Hinge loss in SVM, it is difficult to obtain a faster convergence rate with modern optimization algorithms. In this paper, we introduce two smooth Hinge losses $\psi_G(\alpha;\sigma)$ and $\psi_M(\alpha;\sigma)$ which are infinitely differentiable and converge to the Hinge loss uniformly in $\alpha$ as $\sigma$ tends to $0$. By replacing the Hinge loss with these two smooth Hinge losses, we obtain two smooth support vector machines(SSVMs), respectively. Solving the SSVMs with the Trust Region Newton method (TRON) leads to two quadratically convergent algorithms. Experiments in text classification tasks show that the proposed SSVMs are effective in real-world applications. We also introduce a general smooth convex loss function to unify several commonly-used convex loss functions in machine learning. The general framework provides smooth approximation functions to non-smooth convex loss functions, which can be used to obtain smooth models that can be solved with faster convergent optimization algorithms.

📄 PDF Abstract BibTeX arXiv:2103.00233

Code (0)

등록된 구현이 없습니다.

Tasks

text-classificationText Classification

Methods 이 논문이 사용한 방법론

SVM A Support Vector Machine, or SVM, is a non-parametric supervised learning model. For non-linear classification and regression, they utilise the kernel trick to map inputs…

Similar Papers 제목 키워드 기반

Stochastic smoothing of the top-K calibrated hinge loss for deep imbalanced classification

2022-02-04 · Camille Garcin, Maximilien Servajean, Alexis Joly, Joseph Salmon

In modern classification tasks, the number of labels is getting larger and larger, as is the size of the datasets encountered in practice. As the number of classes increases, class ambiguity and class imbalance become mo…

imbalanced classification

Linear-Core Surrogates: Smooth Loss Functions with Linear Rates for Classification and Structured Prediction

2026-04-30 · Mehryar Mohri, Yutao Zhong arxiv

The choice of loss function in classification involves a fundamental trade-off: smooth losses (like Cross-Entropy) enable fast optimization rates but yield slow square-root consistency bounds, while piecewise-linear loss…

Structured Prediction

The Nyström method for convex loss functions

2020-06-17 · Andrea Della Vecchia, Ernesto de Vito, Jaouad Mourtada, Lorenzo Rosasco

We investigate an extension of classical empirical risk minimization, where the hypothesis space consists of a random subspace within a given Hilbert space. Specifically, we examine the Nystr\"om method where the subspac…

Computational Efficiency

Differentially Private SGD with Non-Smooth Losses

2021-01-22 · Puyu Wang, Yunwen Lei, Yiming Ying, Hai Zhang

In this paper, we are concerned with differentially private {stochastic gradient descent (SGD)} algorithms in the setting of stochastic convex optimization (SCO). Most of the existing work requires the loss to be Lipschi…

Noninteractive Locally Private Learning of Linear Models via Polynomial Approximations

2018-12-17 · Di Wang, Adam Smith, Jinhui Xu

Minimizing a convex risk function is the main step in many basic learning algorithms. We study protocols for convex optimization which provably leak very little about the individual data points that constitute the loss f…