paper-with-me

홈 › Papers

Generalization and Overfitting in Matrix Product State Machine Learning Architectures

2022-08-08 · Artem Strashko, E. Miles Stoudenmire

While overfitting and, more generally, double descent are ubiquitous in machine learning, increasing the number of parameters of the most widely used tensor network, the matrix product state (MPS), has generally lead to monotonic improvement of test performance in previous studies. To better understand the generalization properties of architectures parameterized by MPS, we construct artificial data which can be exactly modeled by an MPS and train the models with different number of parameters. We observe model overfitting for one-dimensional data, but also find that for more complex data overfitting is less significant, while with MNIST image data we do not find any signatures of overfitting. We speculate that generalization properties of MPS depend on the properties of data: with one-dimensional data (for which the MPS ansatz is the most suitable) MPS is prone to overfitting, while with more complex data which cannot be fit by MPS exactly, overfitting may be much less significant.

📄 PDF Abstract BibTeX arXiv:2208.04372

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Test 설명 없음

Similar Papers 제목 키워드 기반

Benign Overfitting with Quantum Kernels

2025-03-21 · Joachim Tomasi, Sandrine Anthoine, Hachem Kadri

Quantum kernels quantify similarity between data points by measuring the inner product between quantum states, computed through quantum circuit measurements. By embedding data into quantum systems, quantum kernel feature…

Detecting overfitting in Neural Networks during long-horizon grokking using Random Matrix Theory

2026-05-12 · Hari K. Prakash, Charles H Martin arxiv

Training Neural Networks (NNs) without overfitting is difficult; detecting that overfitting is difficult as well. We present a novel Random Matrix Theory method that detects the onset of overfitting in deep learning mode…

Benign Overfitting in Out-of-Distribution Generalization of Linear Models

2024-12-19 · Shange Tang, Jiayun Wu, Jianqing Fan, Chi Jin

Benign overfitting refers to the phenomenon where an over-parameterized model fits the training data perfectly, including noise in the data, but still generalizes well to the unseen test data. While prior work provides s…

Out-of-Distribution Generalizationregression

Asymptotic Bayesian Generalization Error in Latent Dirichlet Allocation and Stochastic Matrix Factorization

2017-09-13 · Naoki Hayashi, Sumio Watanabe

Latent Dirichlet allocation (LDA) is useful in document analysis, image processing, and many information systems; however, its generalization performance has been left unknown because it is a singular learning machine to…

Bayesian InferenceTopic Models

A Training-Time Diagnostic for Generalization via the Log-Alignment Ratio

2026-05-27 · Ali Shehper, Ashish Vaswani arxiv

We study the log-alignment ratio (LAR), a measure of parameter-activation alignment, introduced in parameterization theory. We reformulate it as the overlap between a weight spectrum $p$ of the normalized squared singula…