paper-with-me

Papers

Diffusion Models Observe Only Gradients: A Geometric Perspective on Score Matching Errors

2026-06-04 · Naïl B. Khelifa, Richard E. Turner, Ramji Venkataramanan arxiv

Score-based diffusion models are typically trained by minimizing the $L^2$ score matching error, and standard theoretical analyses rely on this quantity to bound the sampling discrepancy between the learned and target distributions. We show the $L^2$ score error is not the right intrinsic measure of marginal distributional quality: a learned diffusion model can incur arbitrarily large $L^2$ score error while perfectly matching the target distribution. By decomposing score errors into a gradient and a solenoidal component (a Helmholtz-Hodge decomposition), we identify the geometric reason behind this: only the gradient component enters the marginal Fokker-Planck dynamics, while the solenoidal component is structurally invisible. We make this precise in three results. First, building on the corrected geometry, we prove an impossibility result: no monotone function of the $L^2$ score error can uniformly lower bound any divergence between the learned and target distributions. Second, we derive an upper bound on the Kullback-Leibler divergence that depends only on the observable gradient component of the error, tightening the standard Girsanov bound for generic score networks, and identifying its looseness as the cost of operating on path-space rather than marginal-space dynamics. Third, we give a tractable estimator of the gradient component via a dual Sobolev identity, which is shown to empirically correlate substantially better with sample quality than the full $L^2$ error.

📄 PDF Abstract BibTeX arXiv:2606.06179

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A Geometric Understanding of Natural Gradient

2022-02-13 · Qinxun Bai, Steven Rosenberg, Wei Xu

While natural gradients have been widely studied from both theoretical and empirical perspectives, we argue that some fundamental theoretical issues regarding the existence of gradients in infinite dimensional function s…

Directional Analysis of Stochastic Gradient Descent via von Mises-Fisher Distributions in Deep learning

2018-09-29 · ICLR 2019 5 · Cheolhyoung Lee, Kyunghyun Cho, Wanmo Kang

Although stochastic gradient descent (SGD) is a driving force behind the recent success of deep learning, our understanding of its dynamics in a high-dimensional parameter space is limited. In recent years, some research…

NeuSD: Surface Completion with Multi-View Text-to-Image Diffusion

2023-12-07 · Savva Ignatyev, Daniil Selikhanovych, Oleg Voynov, Yiqun Wang 외

We present a novel method for 3D surface reconstruction from multiple images where only a part of the object of interest is captured. Our approach builds on two recent developments: surface reconstruction using neural ra…

Surface Reconstruction

DiST-4D: Disentangled Spatiotemporal Diffusion with Metric Depth for 4D Driving Scene Generation

2025-03-19 · Jiazhe Guo, Yikang Ding, Xiwu Chen, Shuo Chen 외

Current generative models struggle to synthesize dynamic 4D driving scenes that simultaneously support temporal extrapolation and spatial novel view synthesis (NVS) without per-scene optimization. A key challenge lies in…

Novel View SynthesisScene Generation

Rethinking Losses for Diffusion Bridge Samplers

2025-06-12 · Sebastian Sanokowski, Lukas Gruber, Christoph Bartmann, Sepp Hochreiter 외

Diffusion bridges are a promising class of deep-learning methods for sampling from unnormalized distributions. Recent works show that the Log Variance (LV) loss consistently outperforms the reverse Kullback-Leibler (rKL)…

Hyperparameter Optimization