paper-with-me

Papers

Consistent Sparse Deep Learning: Theory and Computation

2021-02-25 · Yan Sun, Qifan Song, Faming Liang

Deep learning has been the engine powering many successes of data science. However, the deep neural network (DNN), as the basic model of deep learning, is often excessively over-parameterized, causing many difficulties in training, prediction and interpretation. We propose a frequentist-like method for learning sparse DNNs and justify its consistency under the Bayesian framework: the proposed method could learn a sparse DNN with at most $O(n/\log(n))$ connections and nice theoretical guarantees such as posterior consistency, variable selection consistency and asymptotically optimal generalization bounds. In particular, we establish posterior consistency for the sparse DNN with a mixture Gaussian prior, show that the structure of the sparse DNN can be consistently determined using a Laplace approximation-based marginal posterior inclusion probability approach, and use Bayesian evidence to elicit sparse DNNs learned by an optimization method such as stochastic gradient descent in multiple runs with different initializations. The proposed method is computationally more efficient than standard Bayesian methods for large-scale sparse DNNs. The numerical results indicate that the proposed method can perform very well for large-scale network compression and high-dimensional nonlinear variable selection, both advancing interpretable machine learning.

📄 PDF Abstract BibTeX arXiv:2102.13229

Code (1)

sylydya/Consistent-Sparse-Deep-Learning-Theory-and-Computation 공식 구현 pytorch

Tasks

Deep LearningGeneralization BoundsInterpretable Machine LearningLearning TheoryVariable Selection

Similar Papers 제목 키워드 기반

Random Matrix Theory-guided sparse PCA for single-cell RNA-seq data

2025-09-18 · Victor Chardès arxiv

Single-cell RNA-seq provides detailed molecular snapshots of individual cells but is notoriously noisy. Variability stems from biological differences and technical factors, such as amplification bias and limited RNA capt…

Dimensionality Reduction

An Approach to One-Bit Compressed Sensing Based on Probably Approximately Correct Learning Theory

2017-10-22 · Mehmet Eren Ahsen, Mathukumalli Vidyasagar

In this paper, the problem of one-bit compressed sensing (OBCS) is formulated as a problem in probably approximately correct (PAC) learning. It is shown that the Vapnik-Chervonenkis (VC-) dimension of the set of half-spa…

2kcompressed sensingLearning TheoryPAC learning

Computationally Efficient Robust Estimation of Sparse Functionals

2017-02-24 · Simon S. Du, Sivaraman Balakrishnan, Aarti Singh

Many conventional statistical procedures are extremely sensitive to seemingly minor deviations from modeling assumptions. This problem is exacerbated in modern high-dimensional settings, where the problem dimension can g…

regression

Semantic Folding Theory And its Application in Semantic Fingerprinting

2015-11-28 · Francisco De Sousa Webber

Human language is recognized as a very complex domain since decades. No computer system has been able to reach human levels of performance so far. The only known computational system capable of proper language processing…

Balancing Statistical and Computational Precision: A General Theory and Applications to Sparse Regression

2016-09-23 · Mahsa Taheri, Néhémy Lim, Johannes Lederer

Modern technologies are generating ever-increasing amounts of data. Making use of these data requires methods that are both statistically sound and computationally efficient. Typically, the statistical and computational …

Astronomyfeature selectionregression