paper-with-me

홈 › Papers

An exact information theory of generalization phase transitions in Bayesian diffusion models

2026-07-09 · Henry Hunt, Mason Kamb, Surya Ganguli arxiv

How diffusion models circumvent the curse of dimensionality to learn complex distributions over high dimensional spaces from a finite training set, instead of memorizing it, remains a fundamental mystery. To address this, we introduce analytically tractable Bayesian information restricted diffusion (BIRD) models, in which each pixel observes restricted information about noisy data. A BIRD model time-reverses diffusion by inferring which past training sample produced its current restricted observation using the Bayesian posterior. This model class generalizes existing analytical diffusion models that use spatially local information restriction. We show that spatially local BIRD models closely approximate trained diffusion models \textit{early in training}, across different architectures such as UNets and DiTs. Under minimal assumptions on the data distribution, we identify an information-theoretic phase boundary between memorization and generalization in the joint space of amount of training data, time in the reverse generative process, and amount of information restriction: a BIRD model memorizes when the mutual information between its restricted noisy observations and the training data exceeds the log number of training points, and it generalizes otherwise. Experiments across a range of datasets confirm our theoretically predicted location for the transition. We find that generation proceeds near the edge of memorization: both spatially local BIRD models and early-training diffusion models track the memorization-generalization phase boundary by increasingly restricting information over time. Overall, our results reveal a fundamental role for information restriction in generative AI to circumvent the curse of dimensionality.

📄 PDF Abstract BibTeX arXiv:2607.08041

Code (1)

arxivsub/arXivSub_daily_arxiv ★ 2

Similar Papers 제목 키워드 기반

Exact Phase Transitions in Deep Learning

2022-05-25 · Liu Ziyin, Masahito Ueda

This work reports deep-learning-unique first-order and second-order phase transitions, whose phenomenology closely follows that in statistical physics. In particular, we prove that the competition between prediction erro…

Deep Learning

Causal Inference (C-inf) -- asymmetric scenario of typical phase transitions

2023-01-02 · Agostino Capponi, Mihailo Stojnic

In this paper, we revisit and further explore a mathematically rigorous connection between Causal inference (C-inf) and the Low-rank recovery (LRR) established in [10]. Leveraging the Random duality - Free probability th…

Causal Inference

Phase Transitions for the Information Bottleneck in Representation Learning

2020-01-07 · ICLR 2020 1 · Tailin Wu, Ian Fischer

In the Information Bottleneck (IB), when tuning the relative strength between compression and prediction terms, how do the two terms behave, and what's their relationship with the dataset and the learned representation? …

Representation Learning

Causal Inference (C-inf) -- closed form worst case typical phase transitions

2023-01-02 · Agostino Capponi, Mihailo Stojnic

In this paper we establish a mathematically rigorous connection between Causal inference (C-inf) and the low-rank recovery (LRR). Using Random Duality Theory (RDT) concepts developed in [46,48,50] and novel mathematical …

Causal InferenceForm

Learning in PINNs: Phase transition, total diffusion, and generalization

2024-03-27 · Sokratis J. Anagnostopoulos, Juan Diego Toscano, Nikolaos Stergiopulos, George Em Karniadakis

We investigate the learning dynamics of fully-connected neural networks through the lens of gradient signal-to-noise ratio (SNR), examining the behavior of first-order optimizers like Adam in non-convex objectives. By in…