paper-with-me

홈 › Papers

Diagonal Linear Networks and the Lasso Regularization Path

2025-09-23 · Raphaël Berthier arxiv

Diagonal linear networks are neural networks with linear activation and diagonal weight matrices. Their theoretical interest is that their implicit regularization can be rigorously analyzed: from a small initialization, the training of diagonal linear networks converges to the linear predictor with minimal 1-norm among minimizers of the training loss. In this paper, we deepen this analysis showing that the full training trajectory of diagonal linear networks is closely related to the lasso regularization path. In this connection, the training time plays the role of an inverse regularization parameter. Both rigorous results and simulations are provided to illustrate this conclusion. Under a monotonicity assumption on the lasso regularization path, the connection is exact while in the general case, we show an approximate connection.

📄 PDF Abstract BibTeX arXiv:2509.18766

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A Comparison of Variable Selection Methods for Blockwise Diagonal Designs

2021-09-29 · ICLR 2022 4 · Tracy Ke, Longlin Wang

Lasso is a celebrated method for variable selection in linear models, but it faces challenges when the covariates are moderately or strongly correlated. This motivates alternative approaches such as using a non-convex pe…

Variable Selection

The Well-Tempered Lasso

2018-07-01 · ICML 2018 7 · Yuanzhi Li, Yoram Singer

We study the complexity of the entire regularization path for least squares regression with 1-norm penalty, known as the Lasso. Every regression parameter in the Lasso changes linearly as a function of the regulariz…

regression

The Well Tempered Lasso

2018-06-08 · Yuanzhi Li, Yoram Singer

We study the complexity of the entire regularization path for least squares regression with 1-norm penalty, known as the Lasso. Every regression parameter in the Lasso changes linearly as a function of the regularization…

regression

Lasso Regularization Paths for NARMAX Models via Coordinate Descent

2017-10-02 · Antônio H. Ribeiro, Luis A. Aguirre

We propose a new algorithm for estimating NARMAX models with $L_1$ regularization for models represented as a linear combination of basis functions. Due to the $L_1$-norm penalty the Lasso estimation tends to produce som…

Computational Efficiency

Gaussian Mixture Model with unknown diagonal covariances via continuous sparse regularization

2025-09-16 · Romane Giard, Yohann de Castro, Clément Marteau arxiv

This paper addresses the statistical estimation of Gaussian Mixture Models (GMMs) with unknown diagonal covariances from independent and identically distributed samples. We employ the Beurling-LASSO (BLASSO), a convex op…