paper-with-me

Papers

Preserving Spectral Structure and Statistics in Diffusion Models

2025-12-19 · Baohua Yan, Jennifer Kava, Qingyuan Liu, Xuan Di arxiv

Standard diffusion models (DMs) rely on the total destruction of data into non-informative white noise, forcing the backward process to denoise from a fully unstructured noise state. While ensuring diversity, this results in a cumbersome and computationally intensive image generation task. We address this challenge by proposing new forward and backward process within a mathematically tractable spectral space. Unlike pixel-based DMs, our forward process converges towards an informative Gaussian prior N(mu_hat,Sigma_hat) rather than white noise. Our method, termed Preserving Spectral Structure and Statistics (PreSS) in diffusion models, guides spectral components toward this informative prior while ensuring that corresponding structural signals remain intact at terminal time. This provides a principled starting point for the backward process, enabling high-quality image reconstruction that builds upon preserved spectral structure while maintaining high generative diversity. Experimental results on CIFAR-10, CelebA and CelebA-HQ demonstrate significant reductions in computational complexity, improved visual diversity, less drift, and a smoother diffusion process compared to pixel-based DMs.

📄 PDF Abstract BibTeX arXiv:2512.17873

Code (0)

등록된 구현이 없습니다.

Tasks

Image ReconstructionImage Generation

Similar Papers 제목 키워드 기반

Spectral-Structured Diffusion for Single-Image Rain Removal

2026-03-10 · Yucheng Xing, Xin Wang arxiv

Rain streaks manifest as directional and frequency-concentrated structures that overlap across multiple scales, making single-image rain removal particularly challenging. While diffusion-based restoration models provide …

Computational EfficiencyRain Removal

FreqFormer: Hierarchical Frequency-Domain Attention with Adaptive Spectral Routing for Long-Sequence Video Diffusion Transformers

2026-04-14 · Haopeng Jin arxiv

Long-sequence video diffusion transformers hit a quadratic self-attention cost that dominates runtime and memory for very long token sequences. Most efficient attention methods use one approximation everywhere, yet video…

LGDC: Latent Graph Diffusion via Spectrum-Preserving Coarsening

2025-12-01 · Nagham Osman, Keyue Jiang, Davide Buffelli, Xiaowen Dong 외 arxiv

Graph generation is a critical task across scientific domains. Existing methods fall broadly into two categories: autoregressive models, which iteratively expand graphs, and one-shot models, such as diffusion, which gene…

Graph Generation

GEWDiff: Geometric Enhanced Wavelet-based Diffusion Model for Hyperspectral Image Super-resolution

2025-11-10 · Sirui Wang, Jiang He, Natàlia Blasco Andreo, Xiao Xiang Zhu arxiv

Improving the quality of hyperspectral images (HSIs), such as through super-resolution, is a crucial research area. However, generative modeling for HSIs presents several challenges. Due to their high spectral dimensiona…

Image Super-Resolution

Spectral Progressive Diffusion for Efficient Image and Video Generation

2026-05-18 · Howard Xiao, Brian Chao, Lior Yariv, Gordon Wetzstein arxiv

Diffusion models have been shown to implicitly generate visual content autoregressively in the frequency domain, where low-frequency components are generated earlier in the denoising process while high-frequency details …

Video Generation