paper-with-me

Papers

Maximal Sparsity with Deep Networks?

2016-05-05 · NeurIPS 2016 12 · Bo Xin, Yizhou Wang, Wen Gao, David Wipf

The iterations of many sparse estimation algorithms are comprised of a fixed linear filter cascaded with a thresholding nonlinearity, which collectively resemble a typical neural network layer. Consequently, a lengthy sequence of algorithm iterations can be viewed as a deep network with shared, hand-crafted layer weights. It is therefore quite natural to examine the degree to which a learned network model might act as a viable surrogate for traditional sparse estimation in domains where ample training data is available. While the possibility of a reduced computational budget is readily apparent when a ceiling is imposed on the number of layers, our work primarily focuses on estimation accuracy. In particular, it is well-known that when a signal dictionary has coherent columns, as quantified by a large RIP constant, then most tractable iterative algorithms are unable to find maximally sparse representations. In contrast, we demonstrate both theoretically and empirically the potential for a trained deep network to recover minimal $\ell_0$-norm representations in regimes where existing methods fail. The resulting system is deployed on a practical photometric stereo estimation problem, where the goal is to remove sparse outliers that can disrupt the estimation of surface normals from a 3D scene.

📄 PDF Abstract BibTeX arXiv:1605.01636

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Sparse Signal Estimation by Maximally Sparse Convex Optimization

2013-02-22 · Ivan W. Selesnick, Ilker Bayram

This paper addresses the problem of sparsity penalized least squares for applications in sparse signal processing, e.g. sparse deconvolution. This paper aims to induce sparsity more strongly than L1 norm regularization, …

Grassmannian Packings in Neural Networks: Learning with Maximal Subspace Packings for Diversity and Anti-Sparsity

2019-11-18 · Dian Ang Yap, Nicholas Roberts, Vinay Uday Prabhu

Kernel sparsity ("dying ReLUs") and lack of diversity are commonly observed in CNN kernels, which decreases model capacity. Drawing inspiration from information theory and wireless communications, we demonstrate the inte…

Diversity

Structured Linear CDEs: Maximally Expressive and Parallel-in-Time Sequence Models

2025-05-23 · Benjamin Walker, Lingyi Yang, Nicola Muca Cirone, Cristopher Salvi 외

Structured Linear Controlled Differential Equations (SLiCEs) provide a unifying framework for sequence models with structured, input-dependent state-transition matrices that retain the maximal expressivity of dense matri…

MambaTime Series Classification

Sparse maximal update parameterization: A holistic approach to sparse training dynamics

2024-05-24 · Nolan Dey, Shane Bergsma, Joel Hestness

Several challenges make it difficult for sparse neural networks to compete with dense models. First, setting a large fraction of weights to zero impairs forward and gradient signal propagation. Second, sparse studies oft…

Language ModelingLanguage Modelling

Optimal Transport with Tempered Exponential Measures

2023-09-07 · Ehsan Amid, Frank Nielsen, Richard Nock, Manfred K. Warmuth

In the field of optimal transport, two prominent subfields face each other: (i) unregularized optimal transport, "\`a-la-Kantorovich", which leads to extremely sparse plans but with algorithms that scale poorly, and (ii)…