paper-with-me

Papers

APT: Improving Diffusion Models for High Resolution Image Generation with Adaptive Path Tracing

2025-07-29 · Sangmin Han, Jinho Jeong, Jinwoo Kim, Seon Joo Kim arxiv

Latent Diffusion Models (LDMs) are generally trained at fixed resolutions, limiting their capability when scaling up to high-resolution images. While training-based approaches address this limitation by training on high-resolution datasets, they require large amounts of data and considerable computational resources, making them less practical. Consequently, training-free methods, particularly patch-based approaches, have become a popular alternative. These methods divide an image into patches and fuse the denoising paths of each patch, showing strong performance on high-resolution generation. However, we observe two critical issues for patch-based approaches, which we call `patch-level distribution shift" and `increased patch monotonicity." To address these issues, we propose Adaptive Path Tracing (APT), a framework that combines Statistical Matching to ensure patch distributions remain consistent in upsampled latents and Scale-aware Scheduling to deal with the patch monotonicity. As a result, APT produces clearer and more refined details in high-resolution images. In addition, APT enables a shortcut denoising process, resulting in faster sampling with minimal quality degradation. Our experimental results confirm that APT produces more detailed outputs with improved inference speed, providing a practical approach to high-resolution image generation.

📄 PDF Abstract BibTeX arXiv:2507.21690

Code (0)

등록된 구현이 없습니다.

Tasks

Image Generation

Similar Papers 제목 키워드 기반

Training-free Diffusion Acceleration with Bottleneck Sampling

2025-03-24 · Ye Tian, Xin Xia, Yuxi Ren, Shanchuan Lin 외

Diffusion models have demonstrated remarkable capabilities in visual content generation but remain challenging to deploy due to their high computational cost during inference. This computational burden primarily arises f…

DenoisingImage GenerationVideo Generation

Enhancing Text-to-Image Generation via End-Edge Collaborative Hybrid Super-Resolution

2026-01-21 · Chongbin Yi, Yuxin Liang, Ziqi Zhou, Peng Yang arxiv

Artificial Intelligence-Generated Content (AIGC) has made significant strides, with high-resolution text-to-image (T2I) generation becoming increasingly critical for improving users' Quality of Experience (QoE). Although…

Text-to-Image GenerationImage Enhancement

Foveated Diffusion: Efficient Spatially Adaptive Image and Video Generation

2026-03-24 · Brian Chao, Lior Yariv, Howard Xiao, Gordon Wetzstein arxiv

Diffusion and flow matching models have unlocked unprecedented capabilities for creative content creation, such as interactive image and streaming video generation. The growing demand for higher resolutions, frame rates,…

Video Generation

SupResDiffGAN a new approach for the Super-Resolution task

2025-04-18 · Dawid Kopeć, Wojciech Kozłowski, Maciej Wizerkaniuk, Dawid Krutul 외

In this work, we present SupResDiffGAN, a novel hybrid architecture that combines the strengths of Generative Adversarial Networks (GANs) and diffusion models for super-resolution tasks. By leveraging latent space repres…

Image GenerationSuper-Resolution

High-Resolution Image Editing via Multi-Stage Blended Diffusion

2022-10-24 · Johannes Ackermann, Minjun Li

Diffusion models have shown great results in image generation and in image editing. However, current approaches are limited to low resolutions due to the computational cost of training diffusion models for high-resolutio…

Image GenerationImage InpaintingSuper-ResolutionVocal Bursts Intensity Prediction