paper-with-me

홈 › Papers

OFTSR: One-Step Flow for Image Super-Resolution with Tunable Fidelity-Realism Trade-offs

2024-12-12 · Yuanzhi Zhu, Ruiqing Wang, Shilin Lu, Junnan Li, Hanshu Yan, Kai Zhang

Recent advances in diffusion and flow-based generative models have demonstrated remarkable success in image restoration tasks, achieving superior perceptual quality compared to traditional deep learning approaches. However, these methods either require numerous sampling steps to generate high-quality images, resulting in significant computational overhead, or rely on model distillation, which usually imposes a fixed fidelity-realism trade-off and thus lacks flexibility. In this paper, we introduce OFTSR, a novel flow-based framework for one-step image super-resolution that can produce outputs with tunable levels of fidelity and realism. Our approach first trains a conditional flow-based super-resolution model to serve as a teacher model. We then distill this teacher model by applying a specialized constraint. Specifically, we force the predictions from our one-step student model for same input to lie on the same sampling ODE trajectory of the teacher model. This alignment ensures that the student model's single-step predictions from initial states match the teacher's predictions from a closer intermediate state. Through extensive experiments on challenging datasets including FFHQ (256$\times$256), DIV2K, and ImageNet (256$\times$256), we demonstrate that OFTSR achieves state-of-the-art performance for one-step image super-resolution, while having the ability to flexibly tune the fidelity-realism trade-off. Code and pre-trained models are available at https://github.com/yuanzhi-zhu/OFTSR and https://huggingface.co/Yuanzhi/OFTSR, respectively.

📄 PDF Abstract BibTeX arXiv:2412.09465

Code (1)

yuanzhi-zhu/oftsr 공식 구현 pytorch

Tasks

Image RestorationImage Super-ResolutionSuper-Resolution

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

No more hard prompts: SoftSRV prompting for synthetic data generation

2024-10-21 · Giulia Desalvo, Jean-Fracois Kagy, Lazaros Karydas, Afshin Rostamizadeh 외

We present a novel soft prompt based framework, SoftSRV, that leverages a frozen pre-trained large language model (LLM) to generate targeted synthetic text sequences. Given a sample from the target distribution, our prop…

Language ModelingLanguage ModellingLarge Language ModelMath+1

MFSR: MeanFlow Distillation for One Step Real-World Image Super Resolution

2026-03-21 · Ruiqing Wang, Kai Zhang, Yuanzhi Zhu, Hanshu Yan 외 arxiv

Diffusion- and flow-based models have advanced Real-world Image Super-Resolution (Real-ISR), but their multi-step sampling makes inference slow and hard to deploy. One-step distillation alleviates the cost, yet often deg…

Image Super-Resolution

RFMSR: Residual Flow Matching for Image Super-Resolution

2026-07-14 · Shuwei Huang, Tianyao Luo, Jicheng Liu, Daizong Liu 외 arxiv

Image super-resolution (ISR) has witnessed remarkable progress with diffusion models and flow matching. The dominant text-to-image (T2I) based approaches leverage large-scale foundation models as generative priors, achie…

Image Super-Resolution

Fast Image Super-Resolution via Consistency Rectified Flow

2026-05-12 · Jiaqi Xu, Wenbo Li, Haoze Sun, Fan Li 외 arxiv

Diffusion models (DMs) have demonstrated remarkable success in real-world image super-resolution (SR), yet their reliance on time-consuming multi-step sampling largely hinders their practical applications. While recent e…

Image Super-Resolution

FLowHigh: Towards Efficient and High-Quality Audio Super-Resolution with Single-Step Flow Matching

2025-01-09 · Jun-Hak Yun, Seung-bin Kim, Seong-Whan Lee

Audio super-resolution is challenging owing to its ill-posed nature. Recently, the application of diffusion models in audio super-resolution has shown promising results in alleviating this challenge. However, diffusion-b…

Audio Super-ResolutionComputational EfficiencySpeech EnhancementSuper-Resolution