paper-with-me

Papers

DFU: scale-robust diffusion model for zero-shot super-resolution image generation

2023-11-30 · Alex Havrilla, Kevin Rojas, Wenjing Liao, Molei Tao

Diffusion generative models have achieved remarkable success in generating images with a fixed resolution. However, existing models have limited ability to generalize to different resolutions when training data at those resolutions are not available. Leveraging techniques from operator learning, we present a novel deep-learning architecture, Dual-FNO UNet (DFU), which approximates the score operator by combining both spatial and spectral information at multiple resolutions. Comparisons of DFU to baselines demonstrate its scalability: 1) simultaneously training on multiple resolutions improves FID over training at any single fixed resolution; 2) DFU generalizes beyond its training resolutions, allowing for coherent, high-fidelity generation at higher-resolutions with the same model, i.e. zero-shot super-resolution image-generation; 3) we propose a fine-tuning strategy to further enhance the zero-shot super-resolution image-generation capability of our model, leading to a FID of 11.3 at 1.66 times the maximum training resolution on FFHQ, which no other method can come close to achieving.

📄 PDF Abstract BibTeX arXiv:2401.06144

Code (1)

dahoas/edm 공식 구현 pytorch

Tasks

Image GenerationOperator learningSuper-Resolution

Similar Papers 제목 키워드 기반

Text-guided Explorable Image Super-resolution

2024-03-02 · CVPR 2024 1 · Kanchana Vaishnavi Gandikota, Paramanand Chandramouli

In this paper, we introduce the problem of zero-shot text-guided exploration of the solutions to open-domain image super-resolution. Our goal is to allow users to explore diverse, semantically accurate reconstructions th…

DiversityImage Super-ResolutionSuper-Resolution

Self-Supervised Spatial And Zero-Shot Angular Super-Resolution by Spatial-Angular Implicit Representation For Rotating-View SNR-Efficient Diffusion MRI

2026-05-04 · Yinzhe Wu, Hongyu Rui, Fanwen Wang, Jiahao Huang 외 arxiv

Rotating-view thick-slice acquisition is highly SNR-efficient for mesoscale diffusion MRI (dMRI) but requires numerous rotating views to satisfy Nyquist sampling, resulting in long scan time. We propose a self-supervised…

Zero-shot CT Super-Resolution using Diffusion-based 2D Projection Priors and Signed 3D Gaussians

2025-08-21 · Jeonghyun Noh, Hyun-Jic Oh, Won-Ki Jeong arxiv

Computed tomography (CT) is important in clinical diagnosis, but acquiring high-resolution (HR) CT is constrained by radiation exposure risks. While deep learning-based super-resolution (SR) methods have shown promise fo…

3D Reconstruction

Improved Multi-Shot Diffusion-Weighted MRI with Zero-Shot Self-Supervised Learning Reconstruction

2023-08-09 · Jaejin Cho, Yohan Jun, Xiaoqing Wang, Caique Kobayashi 외

Diffusion MRI is commonly performed using echo-planar imaging (EPI) due to its rapid acquisition time. However, the resolution of diffusion-weighted images is often limited by magnetic field inhomogeneity-related artifac…

Diffusion MRIImage ReconstructionSelf-Supervised Learning

Unlocking Diffusion Hierarchies: Adaptive Timestep Selection for Zero-Shot Segmentation

2026-06-14 · Ramin Nakhli, Mahesh Ramachandran, Luca Ballan arxiv

Zero-shot segmentation has recently shown notable improvement by leveraging the rich visual priors in large-scale text-to-image diffusion models, such as Stable Diffusion. However, current diffusion-based methods often f…