paper-with-me

홈 › Papers

High-Fidelity Two-Step Image Generation via Teacher-Aligned End-to-End Distillation

2026-06-10 · Dongyang Liu, Ruoyi Du, David Liu, Dengyang Jiang, Liangchen Li, Qilong Wu, Zhen Li, Steven C. H. Hoi, Hongsheng Li, Peng Gao arxiv

Few-step diffusion distillation has become increasingly mature for 4-8-step generation, yet pushing further to 2 steps remains challenging. In this work, we introduce Z-Image Turbo++, a high-quality 2-step image generation model distilled from the 8-step Z-Image Turbo teacher. Our method addresses the central bottlenecks of increased task difficulty and limited model capacity in 2-step generation through three simple but effective design choices tailored to this regime. First, we propose Distribution-Aligned Adversarial Learning, which uses teacher-generated images rather than external real images as real samples for GAN training, providing a more attainable and informative adversarial target. Second, we adopt Step-Decoupled Parameterization, assigning independent model parameters to the two denoising steps to better match their distinct capacity demands. Third, we perform End-to-End Training with Iterative Regularization, allowing the first step to receive gradients from final image quality while preserving a meaningful intermediate generation through an explicit step-1 loss. Together, these designs substantially narrow the quality gap between 2-step and 8-step generation in both qualitative and quantitative evaluations, highlighting the potential of carefully tailored distillation strategies for improving the quality-efficiency trade-off in few-step generation.

📄 PDF Abstract BibTeX arXiv:2606.12575

Code (0)

등록된 구현이 없습니다.

Tasks

Image Generation

Similar Papers 제목 키워드 기반

Adversarial Concept Distillation for One-Step Diffusion Personalization

2025-10-23 · Yixiong Yang, Tao Wu, Senmao Li, Shiqi Yang 외 arxiv

Recent progress in accelerating text-to-image diffusion models enables high-fidelity synthesis within a single denoising step. However, customizing the fast one-step models remains challenging, as existing methods consis…

Reward-Aware Trajectory Shaping for Few-step Visual Generation

2026-04-16 · Rui Li, Bingyu Li, Yuanzhi Liang, Haibin Huang 외 arxiv

Achieving high-fidelity generation in extremely few sampling steps has long been a central goal of generative modeling. Existing approaches largely rely on distillation-based frameworks to compress the original multi-ste…

OFTSR: One-Step Flow for Image Super-Resolution with Tunable Fidelity-Realism Trade-offs

2024-12-12 · Yuanzhi Zhu, Ruiqing Wang, Shilin Lu, Junnan Li 외

Recent advances in diffusion and flow-based generative models have demonstrated remarkable success in image restoration tasks, achieving superior perceptual quality compared to traditional deep learning approaches. Howev…

Image RestorationImage Super-ResolutionSuper-Resolution

Data-Forcing Distillation: Restoring Diversity and Fidelity in Few-Step Video Generation

2026-06-16 · Siyi Chen, Shaowei Liu, Yixuan Jia, Zian Wang 외 arxiv

Recent progress has shown promise in distilling multi-step video diffusion models into efficient few-step students. Among them, Distribution Matching Distillation (DMD) and its successor DMD2 achieved strong generation q…

Video Generation

MoGAN: Improving Motion Quality in Video Diffusion via Few-Step Motion Adversarial Post-Training

2025-11-26 · Haotian Xue, Qi Chen, Zhonghao Wang, Xun Huang 외 arxiv

Video diffusion models achieve strong frame-level fidelity but still struggle with motion coherence, dynamics and realism, often producing jitter, ghosting, or implausible dynamics. A key limitation is that the standard …

Video Generation