paper-with-me

Papers

Quality Assessment for AI Generated Images with Instruction Tuning

2024-05-12 · Jiarui Wang, Huiyu Duan, Guangtao Zhai, Xiongkuo Min

Artificial Intelligence Generated Content (AIGC) has grown rapidly in recent years, among which AI-based image generation has gained widespread attention due to its efficient and imaginative image creation ability. However, AI-generated Images (AIGIs) may not satisfy human preferences due to their unique distortions, which highlights the necessity to understand and evaluate human preferences for AIGIs. To this end, in this paper, we first establish a novel Image Quality Assessment (IQA) database for AIGIs, termed AIGCIQA2023+, which provides human visual preference scores and detailed preference explanations from three perspectives including quality, authenticity, and correspondence. Then, based on the constructed AIGCIQA2023+ database, this paper presents a MINT-IQA model to evaluate and explain human preferences for AIGIs from Multi-perspectives with INstruction Tuning. Specifically, the MINT-IQA model first learn and evaluate human preferences for AI-generated Images from multi-perspectives, then via the vision-language instruction tuning strategy, MINT-IQA attains powerful understanding and explanation ability for human visual preference on AIGIs, which can be used for feedback to further improve the assessment capabilities. Extensive experimental results demonstrate that the proposed MINT-IQA model achieves state-of-the-art performance in understanding and evaluating human visual preferences for AIGIs, and the proposed model also achieves competing results on traditional IQA tasks compared with state-of-the-art IQA models. The AIGCIQA2023+ database and MINT-IQA model are available at: https://github.com/IntMeGroup/MINT-IQA.

📄 PDF Abstract BibTeX arXiv:2405.07346

Code (1)

IntMeGroup/MINT-IQA 공식 구현 pytorch

Tasks

Image GenerationImage Quality Assessment

Similar Papers 제목 키워드 기반

ELIQ: A Label-Free Framework for Quality Assessment of Evolving AI-Generated Images

2026-02-03 · Xinyue Li, Zhiming Xu, Min Tang, Zhaolin Cai 외 arxiv

Generative text-to-image models are advancing at an unprecedented pace, continuously shifting the perceptual quality ceiling and rendering previously collected labels unreliable for newer generations. To address this, we…

Picking the Cream of the Crop: Visual-Centric Data Selection with Collaborative Agents

2025-02-27 · Zhenyu Liu, Yunxin Li, Baotian Hu, Wenhan Luo 외

To improve Multimodal Large Language Models' (MLLMs) ability to process images and complex instructions, researchers predominantly curate large-scale visual instruction tuning datasets, which are either sourced from exis…

Image Quality Assessment

ViDA-UGC: Detailed Image Quality Analysis via Visual Distortion Assessment for UGC Images

2025-08-18 · Wenjie Liao, Jieyu Yuan, Yifang Xu, Chunle Guo 외 arxiv

Recent advances in Multimodal Large Language Models (MLLMs) have introduced a paradigm shift for Image Quality Assessment (IQA) from unexplainable image quality scoring to explainable IQA, demonstrating practical applica…

Image Quality AssessmentImage Restoration

Exploring Instruction Data Quality for Explainable Image Quality Assessment

2025-10-04 · Yunhao Li, Sijing Wu, Huiyu Duan, Yucheng Zhu 외 arxiv

In recent years, with the rapid development of large multimodal models (LMMs), explainable image quality assessment (IQA) has attracted increasing attention, aiming to understand the perceptual quality problems of images…

Image Quality Assessment

Better Reasoning with Less Data: Enhancing VLMs Through Unified Modality Scoring

2025-06-10 · Mingjie Xu, Andrew Estornell, Hongzheng Yang, Yuzhi Zhao 외

The application of visual instruction tuning and other post-training techniques has significantly enhanced the capabilities of Large Language Models (LLMs) in visual understanding, enriching Vision-Language Models (VLMs)…

Image Captioning