paper-with-me

Papers

Does Learning Require Memorization? A Short Tale about a Long Tail

2019-06-12 · Vitaly Feldman

State-of-the-art results on image recognition tasks are achieved using over-parameterized learning algorithms that (nearly) perfectly fit the training set and are known to fit well even random labels. This tendency to memorize the labels of the training data is not explained by existing theoretical analyses. Memorization of the training data also presents significant privacy risks when the training data contains sensitive personal information and thus it is important to understand whether such memorization is necessary for accurate learning. We provide the first conceptual explanation and a theoretical model for this phenomenon. Specifically, we demonstrate that for natural data distributions memorization of labels is necessary for achieving close-to-optimal generalization error. Crucially, even labels of outliers and noisy labels need to be memorized. The model is motivated and supported by the results of several recent empirical works. In our model, data is sampled from a mixture of subpopulations and our results show that memorization is necessary whenever the distribution of subpopulation frequencies is long-tailed. Image and text data is known to be long-tailed and therefore our results establish a formal link between these empirical phenomena. Our results allow to quantify the cost of limiting memorization in learning and explain the disparate effects that privacy and model compression have on different subgroups.

📄 PDF Abstract BibTeX arXiv:1906.05271

Code (0)

등록된 구현이 없습니다.

Tasks

MemorizationModel Compression

Similar Papers 제목 키워드 기반

Representation in large language models

2025-01-01 · Cameron C. Yetman

The extraordinary success of recent Large Language Models (LLMs) on a diverse array of tasks has led to an explosion of scientific and philosophical theorizing aimed at explaining how they do what they do. Unfortunately,…

Memorization

Provable Separations between Memorization and Generalization in Diffusion Models

2025-11-05 · Zeqi Ye, Qijie Zhu, Molei Tao, Minshuo Chen arxiv

Diffusion models have achieved remarkable success across diverse domains, but they remain vulnerable to memorization -- reproducing training data rather than generating novel outputs. This not only limits their creative …

Modifying Memories in Transformer Models

2020-12-01 · Chen Zhu, Ankit Singh Rawat, Manzil Zaheer, Srinadh Bhojanapalli 외

Large Transformer models have achieved impressive performance in many natural language tasks. In particular, Transformer based language models have been shown to have great capabilities in encoding factual knowledge in t…

Memorization

Tackling Intertwined Data and Device Heterogeneities in Federated Learning with Unlimited Staleness

2023-09-24 · Haoming Wang, Wei Gao

Federated Learning (FL) can be affected by data and device heterogeneities, caused by clients' different local data distributions and latencies in uploading model updates (i.e., staleness). Traditional schemes consider t…

Computational EfficiencyFederated Learning

A Tale of Two Animats: What does it take to have goals?

2017-05-30 · Larissa Albantakis

What does it take for a system, biological or not, to have goals? Here, this question is approached in the context of in silico artificial evolution. By examining the informational and causal properties of artificial org…