paper-with-me

Papers

Tighter Expected Generalization Error Bounds via Convexity of Information Measures

2022-02-24 · Gholamali Aminian, Yuheng Bu, Gregory Wornell, Miguel Rodrigues

Generalization error bounds are essential to understanding machine learning algorithms. This paper presents novel expected generalization error upper bounds based on the average joint distribution between the output hypothesis and each input training sample. Multiple generalization error upper bounds based on different information measures are provided, including Wasserstein distance, total variation distance, KL divergence, and Jensen-Shannon divergence. Due to the convexity of the information measures, the proposed bounds in terms of Wasserstein distance and total variation distance are shown to be tighter than their counterparts based on individual samples in the literature. An example is provided to demonstrate the tightness of the proposed generalization error bounds.

📄 PDF Abstract BibTeX arXiv:2202.12150

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Tighter expected generalization error bounds via Wasserstein distance

2021-01-22 · NeurIPS 2021 12 · Borja Rodríguez-Gálvez, Germán Bassi, Ragnar Thobaben, Mikael Skoglund

This work presents several expected generalization error bounds based on the Wasserstein distance. More specifically, it introduces full-dataset, single-letter, and random-subset bounds, and their analogues in the random…

Class-wise Generalization Error: an Information-Theoretic Analysis

2024-01-05 · Firas Laakom, Yuheng Bu, Moncef Gabbouj

Existing generalization theories of supervised learning typically take a holistic approach and provide bounds for the expected generalization over the whole data distribution, which implicitly assumes that the model gene…

Generalization Bounds

Learning Algorithm Generalization Error Bounds via Auxiliary Distributions

2022-10-02 · Gholamali Aminian, Saeed Masiha, Laura Toni, Miguel R. D. Rodrigues

Generalization error bounds are essential for comprehending how well machine learning models work. In this work, we suggest a novel method, i.e., the Auxiliary Distribution Method, that leads to new upper bounds on expec…

Understanding the Generalization Ability of Deep Learning Algorithms: A Kernelized Renyi's Entropy Perspective

2023-05-02 · Yuxin Dong, Tieliang Gong, Hong Chen, Chen Li

Recently, information theoretic analysis has become a popular framework for understanding the generalization behavior of deep neural networks. It allows a direct analysis for stochastic gradient/Langevin descent (SGD/SGL…

On generalization bounds for deep networks based on loss surface implicit regularization

2022-01-12 · Masaaki Imaizumi, Johannes Schmidt-Hieber

The classical statistical learning theory implies that fitting too many parameters leads to overfitting and poor performance. That modern deep neural networks generalize well despite a large number of parameters contradi…

Generalization BoundsLearning Theory