paper-with-me

홈 › Papers

Q-Refine: A Perceptual Quality Refiner for AI-Generated Image

2024-01-02 · Chunyi Li, HaoNing Wu, ZiCheng Zhang, Hongkun Hao, Kaiwei Zhang, Lei Bai, Xiaohong Liu, Xiongkuo Min, Weisi Lin, Guangtao Zhai

With the rapid evolution of the Text-to-Image (T2I) model in recent years, their unsatisfactory generation result has become a challenge. However, uniformly refining AI-Generated Images (AIGIs) of different qualities not only limited optimization capabilities for low-quality AIGIs but also brought negative optimization to high-quality AIGIs. To address this issue, a quality-award refiner named Q-Refine is proposed. Based on the preference of the Human Visual System (HVS), Q-Refine uses the Image Quality Assessment (IQA) metric to guide the refining process for the first time, and modify images of different qualities through three adaptive pipelines. Experimental shows that for mainstream T2I models, Q-Refine can perform effective optimization to AIGIs of different qualities. It can be a general refiner to optimize AIGIs from both fidelity and aesthetic quality levels, thus expanding the application of the T2I generation models.

📄 PDF Abstract BibTeX arXiv:2401.01117

Code (1)

q-future/q-refine 공식 구현 pytorch

Tasks

Image Quality Assessment

Similar Papers 제목 키워드 기반

Diffiner: A Versatile Diffusion-based Generative Refiner for Speech Enhancement

2022-10-27 · Ryosuke Sawata, Naoki Murata, Yuhta Takida, Toshimitsu Uesaka 외

Although deep neural network (DNN)-based speech enhancement (SE) methods outperform the previous non-DNN-based ones, they often degrade the perceptual quality of generated outputs. To tackle this problem, we introduce a …

DenoisingSpeech Enhancement

SpeechRefiner: Towards Perceptual Quality Refinement for Front-End Algorithms

2025-06-16 · Sirui Li, Shuai Wang, Zhijun Liu, Zhongjie Jiang 외

Speech pre-processing techniques such as denoising, de-reverberation, and separation, are commonly employed as front-ends for various downstream speech processing tasks. However, these methods can sometimes be inadequate…

Denoising

DIR-TIR: Dialog-Iterative Refinement for Text-to-Image Retrieval

2025-11-18 · Zongwei Zhen, Biqing Zeng arxiv

This paper addresses the task of interactive, conversational text-to-image retrieval. Our DIR-TIR framework progressively refines the target image search through two specialized modules: the Dialog Refiner Module and the…

Image Retrieval

Diffusion-based Signal Refiner for Speech Separation

2023-05-10 · Masato Hirano, Kazuki Shimada, Yuichiro Koyama, Shusuke Takahashi 외

We have developed a diffusion-based speech refiner that improves the reference-free perceptual quality of the audio predicted by preceding single-channel speech separation models. Although modern deep neural network-base…

DenoisingSpeech EnhancementSpeech Separation

EditRefiner: A Human-Aligned Agentic Framework for Image Editing Refinement

2026-05-08 · Zitong Xu, Huiyu Duan, Yifei Nie, Mingda Du 외 arxiv

Recent text-guided image editing (TIE) models have made remarkable progress, yet edited images still frequently suffer from fine-grained issues such as unnatural objects, lighting mismatch, and unexpected changes. Existi…

Instruction FollowingImage Editing