paper-with-me

홈 › Papers

On information captured by neural networks: connections with memorization and generalization

2023-06-28 · Hrayr Harutyunyan

Despite the popularity and success of deep learning, there is limited understanding of when, how, and why neural networks generalize to unseen examples. Since learning can be seen as extracting information from data, we formally study information captured by neural networks during training. Specifically, we start with viewing learning in presence of noisy labels from an information-theoretic perspective and derive a learning algorithm that limits label noise information in weights. We then define a notion of unique information that an individual sample provides to the training of a deep network, shedding some light on the behavior of neural networks on examples that are atypical, ambiguous, or belong to underrepresented subpopulations. We relate example informativeness to generalization by deriving nonvacuous generalization gap bounds. Finally, by studying knowledge distillation, we highlight the important role of data and label complexity in generalization. Overall, our findings contribute to a deeper understanding of the mechanisms underlying neural network generalization.

📄 PDF Abstract BibTeX arXiv:2306.15918

Code (1)

awslabs/aws-cv-unique-information 공식 구현 pytorch

Tasks

InformativenessKnowledge DistillationMemorization

Similar Papers 제목 키워드 기반

Learning and Memorization

2018-07-01 · ICML 2018 7 · Satrajit Chatterjee

In the machine learning research community, it is generally believed that there is a tension between memorization and generalization. In this work we examine to what extent this tension exists by exploring if it is …

Memorization

A Closer Look at Memorization in Deep Networks

2017-06-16 · ICML 2017 8 · Devansh Arpit, Stanisław Jastrzębski, Nicolas Ballas, David Krueger 외

We examine the role of memorization in deep learning, drawing connections to capacity, generalization, and adversarial robustness. While deep networks are capable of memorizing noise data, our results suggest that they t…

Adversarial RobustnessMemorization

On the Privacy Effect of Data Enhancement via the Lens of Memorization

2022-08-17 · Xiao Li, Qiongxiu Li, Zhanhao Hu, Xiaolin Hu

Machine learning poses severe privacy concerns as it has been shown that the learned models can reveal sensitive information about their training data. Many works have investigated the effect of widely adopted data augme…

Adversarial RobustnessData AugmentationMemorization

Captured by Captions: On Memorization and its Mitigation in CLIP Models

2025-02-11 · Wenhao Wang, Adam Dziedzic, Grace C. Kim, Michael Backes 외

Multi-modal models, such as CLIP, have demonstrated strong performance in aligning visual and textual representations, excelling in tasks like image retrieval and zero-shot classification. Despite this success, the mecha…

Image RetrievalMemorizationSelf-Supervised Learningzero-shot-classification+1

How much do language models memorize?

2025-05-30 · John X. Morris, Chawin Sitawarin, Chuan Guo, Narine Kokhlikyan 외

We propose a new method for estimating how much a model ``knows'' about a datapoint and use it to measure the capacity of modern language models. Prior studies of language model memorization have struggled to disentangle…

Language ModelingLanguage ModellingMemorization