paper-with-me

홈 › Papers

Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model

2025-06-05 · Zelu Qi, Ping Shi, Chaoyang Zhang, Shuqi Wang, Fei Zhao, Da Pan, Zefeng Ying

The development of AI-Generated Video (AIGV) technology has been remarkable in recent years, significantly transforming the paradigm of video content production. However, AIGVs still suffer from noticeable visual quality defects, such as noise, blurriness, frame jitter and low dynamic degree, which severely impact the user's viewing experience. Therefore, an effective automatic visual quality assessment is of great importance for AIGV content regulation and generative model improvement. In this work, we decompose the visual quality of AIGVs into three dimensions: technical quality, motion quality, and video semantics. For each dimension, we design corresponding encoder to achieve effective feature representation. Moreover, considering the outstanding performance of large language models (LLMs) in various vision and language tasks, we introduce a LLM as the quality regression module. To better enable the LLM to establish reasoning associations between multi-dimensional features and visual quality, we propose a specially designed multi-modal prompt engineering framework. Additionally, we incorporate LoRA fine-tuning technology during the training phase, allowing the LLM to better adapt to specific tasks. Our proposed method achieved \textbf{second place} in the NTIRE 2025 Quality Assessment of AI-Generated Content Challenge: Track 2 AI Generated video, demonstrating its effectiveness. Codes can be obtained at https://github.com/QiZelu/AIGVEval.

📄 PDF Abstract BibTeX arXiv:2506.04715

Code (1)

qizelu/aigveval 공식 구현 pytorch

Tasks

Prompt Engineering

Similar Papers 제목 키워드 기반

Exploring AIGC Video Quality: A Focus on Visual Harmony, Video-Text Consistency and Domain Distribution Gap

2024-04-21 · Bowen Qu, Xiaoyu Liang, Shangkun Sun, Wei Gao

The recent advancements in Text-to-Video Artificial Intelligence Generated Content (AIGC) have been remarkable. Compared with traditional videos, the assessment of AIGC videos encounters various challenges: visual incons…

Common Sense Reasoning

VideoAesBench: Benchmarking the Video Aesthetics Perception Capabilities of Large Multimodal Models

2026-01-29 · Yunhao Li, Sijing Wu, Zhilin Gao, Zicheng Zhang 외 arxiv

Large multimodal models (LMMs) have demonstrated outstanding capabilities in various visual perception tasks, which has in turn made the evaluation of LMMs significant. However, the capability of video aesthetic quality …

LoViF 2026 The First Challenge on Holistic Quality Assessment for 4D World Model (PhyScore)

2026-05-06 · Wei Luo, Yiting Lu, Xin Li, Haoran Li 외 arxiv

This paper reports on the LoViF 2026 PhyScore challenge, a competition on holistic quality assessment of world-model-generated videos across both 2D and 4D generation settings. The challenge is motivated by a central gap…

Video Alignment

Benchmarking Multi-dimensional AIGC Video Quality Assessment: A Dataset and Unified Model

2024-07-31 · Zhichao Zhang, Wei Sun, Xinyue Li, Jun Jia 외

In recent years, artificial intelligence (AI)-driven video generation has gained significant attention. Consequently, there is a growing need for accurate video quality assessment (VQA) metrics to evaluate the perceptual…

BenchmarkingLarge Language ModelVideo AlignmentVideo Generation+2

Perceptual Quality Assessment of Omnidirectional Audio-visual Signals

2023-07-20 · Xilei Zhu, Huiyu Duan, Yuqin Cao, Yuxin Zhu 외

Omnidirectional videos (ODVs) play an increasingly important role in the application fields of medical, education, advertising, tourism, etc. Assessing the quality of ODVs is significant for service-providers to improve …