paper-with-me

Papers

Compression Implies Generalization

2021-06-15 · Allan Grønlund, Mikael Høgsgaard, Lior Kamma, Kasper Green Larsen

Explaining the surprising generalization performance of deep neural networks is an active and important line of research in theoretical machine learning. Influential work by Arora et al. (ICML'18) showed that, noise stability properties of deep nets occurring in practice can be used to provably compress model representations. They then argued that the small representations of compressed networks imply good generalization performance albeit only of the compressed nets. Extending their compression framework to yield generalization bounds for the original uncompressed networks remains elusive. Our main contribution is the establishment of a compression-based framework for proving generalization bounds. The framework is simple and powerful enough to extend the generalization bounds by Arora et al. to also hold for the original network. To demonstrate the flexibility of the framework, we also show that it allows us to give simple proofs of the strongest known generalization bounds for other popular machine learning models, namely Support Vector Machines and Boosting.

📄 PDF Abstract BibTeX arXiv:2106.07989

Code (0)

등록된 구현이 없습니다.

Tasks

BIG-bench Machine LearningGeneralization Bounds

Similar Papers 제목 키워드 기반

Non-Vacuous Generalization Bounds at the ImageNet Scale: A PAC-Bayesian Compression Approach

2018-04-16 · ICLR 2019 5 · Wenda Zhou, Victor Veitch, Morgane Austern, Ryan P. Adams 외

Modern neural networks are highly overparameterized, with capacity to substantially overfit to training data. Nevertheless, these networks often generalize well in practice. It has also been observed that trained network…

Generalization Bounds

Reasoning About Generalization via Conditional Mutual Information

2020-01-24 · Thomas Steinke, Lydia Zakynthinou

We provide an information-theoretic framework for studying the generalization properties of machine learning algorithms. Our framework ties together existing approaches, including uniform convergence bounds and recent me…

BIG-bench Machine Learning

Geometric and Information Compression of Representations in Deep Learning

2026-06-19 · Linara Adilova, Henning Petzka, Asja Fischer, Bernhard C. Geiger arxiv

Deep neural networks transform input data into latent representations that support a wide range of downstream tasks. These representations can be characterized along information-theoretic and geometric dimensions, but th…

Unlabeled sample compression schemes and corner peelings for ample and maximum classes

2018-12-05 · Jérémie Chalopin, Victor Chepoi, Shay Moran, Manfred K. Warmuth

We examine connections between combinatorial notions that arise in machine learning and topological notions in cubical/simplicial geometry. These connections enable to export results from geometry to machine learning. Ou…

BIG-bench Machine Learning

High-arity Sample Compression

2026-05-12 · Leonardo N. Coregliano, William Opich arxiv

Recently, a series of works have started studying variations of concepts from learning theory for product spaces, which can be collected under the name high-arity learning theory. In this work, we consider a high-arity v…