paper-with-me

Papers

Optimizing Non-decomposable Measures with Deep Networks

2018-01-31 · Amartya Sanyal, Pawan Kumar, Purushottam Kar, Sanjay Chawla, Fabrizio Sebastiani

We present a class of algorithms capable of directly training deep neural networks with respect to large families of task-specific performance measures such as the F-measure and the Kullback-Leibler divergence that are structured and non-decomposable. This presents a departure from standard deep learning techniques that typically use squared or cross-entropy loss functions (that are decomposable) to train neural networks. We demonstrate that directly training with task-specific loss functions yields much faster and more stable convergence across problems and datasets. Our proposed algorithms and implementations have several novel features including (i) convergence to first order stationary points despite optimizing complex objective functions; (ii) use of fewer training samples to achieve a desired level of convergence, (iii) a substantial reduction in training time, and (iv) a seamless integration of our implementation into existing symbolic gradient frameworks. We implement our techniques on a variety of deep architectures including multi-layer perceptrons and recurrent neural networks and show that on a variety of benchmark and real data sets, our algorithms outperform traditional approaches to training deep networks, as well as some recent approaches to task-specific training of neural networks.

📄 PDF Abstract BibTeX arXiv:1802.00086

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Optimizing Non-decomposable Performance Measures: A Tale of Two Classes

2015-05-26 · Harikrishna Narasimhan, Purushottam Kar, Prateek Jain

Modern classification problems frequently present mild to severe label imbalance as well as specific requirements on classification characteristics, and require optimizing performance measures that are non-decomposable o…

General ClassificationVocal Bursts Valence Prediction

Cost-Sensitive Self-Training for Optimizing Non-Decomposable Metrics

2023-04-28 · Harsh Rangwani, Shrinivas Ramasubramanian, Sho Takemori, Kato Takashi 외

Self-training based semi-supervised learning algorithms have enabled the learning of highly accurate deep neural networks, using only a fraction of labeled data. However, the majority of work on self-training has focused…

Variance Reduced Stochastic Proximal Algorithm for AUC Maximization

2019-11-08 · Soham Dan, Dushyant Sahoo

Stochastic Gradient Descent has been widely studied with classification accuracy as a performance measure. However, these stochastic algorithms cannot be directly used when non-decomposable pairwise performance measures …

Decomposable sums and their implications on naturally quasiconvex risk measures

2022-01-14 · Çağın Ararat, Barış Bilir, Elisa Mastrogiacomo

Convexity and quasiconvexity are two properties that capture the concept of diversification for risk measures. Between the two, there is natural quasiconvexity, an old but not so well-known property weaker than convexity…

Algorithmic Foundations of Empirical X-risk Minimization

2022-06-01 · Tianbao Yang

This manuscript introduces a new optimization framework for machine learning and AI, named {\bf empirical X-risk minimization (EXM)}. X-risk is a term introduced to represent a family of compositional measures or objecti…

Bilevel OptimizationInformation RetrievalRetrieval