paper-with-me

홈 › Papers

Information bottleneck theory of high-dimensional regression: relevancy, efficiency and optimality

2022-08-08 · Vudtiwat Ngampruetikorn, David J. Schwab

Avoiding overfitting is a central challenge in machine learning, yet many large neural networks readily achieve zero training loss. This puzzling contradiction necessitates new approaches to the study of overfitting. Here we quantify overfitting via residual information, defined as the bits in fitted models that encode noise in training data. Information efficient learning algorithms minimize residual information while maximizing the relevant bits, which are predictive of the unknown generative models. We solve this optimization to obtain the information content of optimal algorithms for a linear regression problem and compare it to that of randomized ridge regression. Our results demonstrate the fundamental trade-off between residual and relevant information and characterize the relative information efficiency of randomized regression with respect to optimal algorithms. Finally, using results from random matrix theory, we reveal the information complexity of learning a linear map in high dimensions and unveil information-theoretic analogs of double and multiple descent phenomena.

📄 PDF Abstract BibTeX arXiv:2208.03848

Code (0)

등록된 구현이 없습니다.

Tasks

regressionVocal Bursts Intensity Prediction

Methods 이 논문이 사용한 방법론

Linear Regression Linear Regression is a method for modelling a relationship between a dependent variable and independent variables. These models can be fit with numerous approaches. The most…

Similar Papers 제목 키워드 기반

Bridging factor and sparse models

2021-02-22 · Jianqing Fan, Ricardo Masini, Marcelo C. Medeiros

Factor and sparse models are two widely used methods to impose a low-dimensional structure in high-dimensions. However, they are seemingly mutually exclusive. We propose a lifting method that combines the merits of these…

Model SelectionregressionTime Series Analysis

High Dimensional Time Series Regression Models: Applications to Statistical Learning Methods

2023-08-27 · Christis Katsouris

These lecture notes provide an overview of existing methodologies and recent developments for estimation and inference with high dimensional time series regression models. First, we present main limit theory results for …

regressionTime SeriesTime Series AnalysisTime Series Regression

Learning to Compress: Local Rank and Information Compression in Deep Neural Networks

2024-10-10 · Niket Patel, Ravid Shwartz-Ziv

Deep neural networks tend to exhibit a bias toward low-rank solutions during training, implicitly learning low-dimensional feature representations. This paper investigates how deep multilayer perceptrons (MLPs) encode th…

Representation Learning

On Universal Features for High-Dimensional Learning and Inference

2019-11-20 · Shao-Lun Huang, Anuran Makur, Gregory W. Wornell, Lizhong Zheng

We consider the problem of identifying universal low-dimensional features from high-dimensional data for inference tasks in settings involving learning. For such problems, we introduce natural notions of universality and…

Collaborative FilteringregressionVocal Bursts Intensity Prediction

On the Information Plane of Autoencoders

2020-05-15 · Nicolás I. Tapia, Pablo A. Estévez

The training dynamics of hidden layers in deep learning are poorly understood in theory. Recently, the Information Plane (IP) was proposed to analyze them, which is based on the information-theoretic concept of mutual in…

Information PlaneMutual Information Estimation