paper-with-me

홈 › Papers

Detecting Memorization in ReLU Networks

2018-10-08 · ICLR 2019 5 · Edo Collins, Siavash Arjomand Bigdeli, Sabine Süsstrunk

We propose a new notion of `non-linearity' of a network layer with respect to an input batch that is based on its proximity to a linear system, which is reflected in the non-negative rank of the activation matrix. We measure this non-linearity by applying non-negative factorization to the activation matrix. Considering batches of similar samples, we find that high non-linearity in deep layers is indicative of memorization. Furthermore, by applying our approach layer-by-layer, we find that the mechanism for memorization consists of distinct phases. We perform experiments on fully-connected and convolutional neural networks trained on several image and audio datasets. Our results demonstrate that as an indicator for memorization, our technique can be used to perform early stopping.

📄 PDF Abstract BibTeX arXiv:1810.03372

Code (0)

등록된 구현이 없습니다.

Tasks

Memorization

Similar Papers 제목 키워드 기반

The Cost of Robustness: Tighter Bounds on Parameter Complexity for Robust Memorization in ReLU Nets

2025-10-28 · Yujun Kim, Chaewon Moon, Chulhee Yun arxiv

We study the parameter complexity of robust memorization for $\mathrm{ReLU}$ networks: the number of parameters required to interpolate any given dataset with $ε$-separation between differently labeled points, while ensu…

Small ReLU networks are powerful memorizers: a tight analysis of memorization capacity

2018-10-17 · NeurIPS 2019 12 · Chulhee Yun, Suvrit Sra, Ali Jadbabaie

We study finite sample expressivity, i.e., memorization power of ReLU networks. Recent results require $N$ hidden nodes to memorize/interpolate arbitrary $N$ data points. In contrast, by exploiting depth, we show that 3-…

Memorization

Memorization capacity of deep ReLU neural networks characterized by width and depth

2026-03-10 · Xin Yang, Yunfei Yang arxiv

This paper studies the memorization capacity of deep neural networks with ReLU activation. Specifically, we investigate the minimal size of such networks to memorize any $N$ data points in the unit ball with pairwise sep…

Generalization of Diffusion Models Arises with a Balanced Representation Space

2025-12-24 · Zekai Zhang, Xiao Li, Xiang Li, Lianghe Shi 외 arxiv

Diffusion models excel at generating high-quality, diverse samples, yet they risk memorizing training data when overfit to the training objective. We analyze the distinctions between memorization and generalization in di…

Representation Learning

On the Optimal Memorization Power of ReLU Neural Networks

2021-10-07 · ICLR 2022 4 · Gal Vardi, Gilad Yehudai, Ohad Shamir

We study the memorization power of feedforward ReLU neural networks. We show that such networks can memorize any $N$ points that satisfy a mild separability assumption using $\tilde{O}\left(\sqrt{N}\right)$ parameters. K…

Memorization