paper-with-me

Papers

Path Regularization: A Convexity and Sparsity Inducing Regularization for Parallel ReLU Networks

2021-10-18 · NeurIPS 2023 11

Understanding the fundamental principles behind the success of deep neural networks is one of the most important open questions in the current literature. To this end, we study the training problem of deep neural networks and introduce an analytic approach to unveil hidden convexity in the optimization landscape. We consider a deep parallel ReLU network architecture, which also includes standard deep networks and ResNets as its special cases. We then show that pathwise regularized training problems can be represented as an exact convex optimization problem. We further prove that the equivalent convex problem is regularized via a group sparsity inducing norm. Thus, a path regularized parallel ReLU network can be viewed as a parsimonious convex model in high dimensions. More importantly, since the original training problem may not be trainable in polynomial-time, we propose an approximate algorithm with a fully polynomial-time complexity in all data dimensions. Then, we prove strong global optimality guarantees for this algorithm. We also provide experiments corroborating our theory.

📄 PDF Abstract BibTeX arXiv:2110.09548

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Sparse Signal Estimation by Maximally Sparse Convex Optimization

2013-02-22 · Ivan W. Selesnick, Ilker Bayram

This paper addresses the problem of sparsity penalized least squares for applications in sparse signal processing, e.g. sparse deconvolution. This paper aims to induce sparsity more strongly than L1 norm regularization, …

Nonconvex Sparse Logistic Regression with Weakly Convex Regularization

2017-08-07 · Xinyue Shen, Yuantao Gu

In this work we propose to fit a sparse logistic regression model by a weakly convex regularized nonconvex optimization problem. The idea is based on the finding that a weakly convex function as an approximation of the $…

regression

Regularization can make diffusion models more efficient

2025-02-13 · Mahsa Taheri, Johannes Lederer

Diffusion models are one of the key architectures of generative AI. Their main drawback, however, is the computational costs. This study indicates that the concept of sparsity, well known especially in statistics, can pr…

Nonconvex Regularization for Feature Selection in Reinforcement Learning

2025-09-19 · Kyohei Suzuki, Konstantinos Slavakis arxiv

This work proposes an efficient batch algorithm for feature selection in reinforcement learning (RL) with theoretical convergence guarantees. To mitigate the estimation bias inherent in conventional regularization scheme…

Reinforcement Learning

Regularization vs. Relaxation: A conic optimization perspective of statistical variable selection

2015-10-20 · Hongbo Dong, Kun Chen, Jeff Linderoth

Variable selection is a fundamental task in statistical data analysis. Sparsity-inducing regularization methods are a popular class of methods that simultaneously perform variable selection and model estimation. The cent…

Combinatorial OptimizationVariable Selection