paper-with-me

홈 › Papers

Deep Video Quality Assessor: From Spatio-temporal Visual Sensitivity to A Convolutional Neural Aggregation Network

2018-09-01 · ECCV 2018 9 · Woojae Kim, Jongyoo Kim, Sewoong Ahn, Jinwoo Kim, Sang-Hoon Lee

Incorporating spatio-temporal human visual perception into video quality assessment (VQA) remains a formidable issue. Previous statistical or computational models of spatio-temporal perception have limitations to be applied to the general VQA algorithms. In this paper, we propose a novel full-reference (FR) VQA framework named Deep Video Quality Assessor (DeepVQA) to quantify the spatio-temporal visual perception via a convolutional neural network (CNN) and a convolutional neural aggregation network (CNAN). Our framework enables to figure out the spatio-temporal sensitivity behavior through learning in accordance with the subjective score. In addition, to manipulate the temporal variation of distortions, we propose a novel temporal pooling method using an attention model. In the experiment, we show DeepVQA remarkably achieves the state-of-the-art prediction accuracy of more than 0.9 correlation, which is ~5% higher than those of conventional methods on the LIVE and CSIQ video databases.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

SensitivityVideo Quality AssessmentVisual Question Answering (VQA)

Similar Papers 제목 키워드 기반

AIGV-Assessor: Benchmarking and Evaluating the Perceptual Quality of Text-to-Video Generation with LMM

2024-11-26 · CVPR 2025 1 · Jiarui Wang, Huiyu Duan, Guangtao Zhai, Juntong Wang 외

The rapid advancement of large multimodal models (LMMs) has led to the rapid expansion of artificial intelligence generated videos (AIGVs), which highlights the pressing need for effective video quality assessment (VQA) …

BenchmarkingText-to-Video GenerationVideo GenerationVideo Quality Assessment+1

TDVE-Assessor: Benchmarking and Evaluating the Quality of Text-Driven Video Editing with LMMs

2025-05-26 · Juntong Wang, Jiarui Wang, Huiyu Duan, Guangtao Zhai 외

Text-driven video editing is rapidly advancing, yet its rigorous evaluation remains challenging due to the absence of dedicated video quality assessment (VQA) models capable of discerning the nuances of editing quality. …

BenchmarkingLarge Language ModelVideo EditingVideo Quality Assessment+1

Saliency-Aware Spatio-Temporal Artifact Detection for Compressed Video Quality Assessment

2023-01-03 · Liqun Lin, Yang Zheng, Weiling Chen, Chengdong Lan 외

Compressed videos often exhibit visually annoying artifacts, known as Perceivable Encoding Artifacts (PEAs), which dramatically degrade video visual quality. Subjective and objective measures capable of identifying and q…

Artifact DetectionBlockingVideo Quality Assessment

R-AVST: Empowering Video-LLMs with Fine-Grained Spatio-Temporal Reasoning in Complex Audio-Visual Scenarios

2025-11-21 · Lu Zhu, Tiantian Geng, Yangye Chen, Teng Wang 외 arxiv

Recently, rapid advancements have been made in multimodal large language models (MLLMs), especially in video understanding tasks. However, current research focuses on simple video scenarios, failing to reflect the comple…

Reinforcement LearningVisual Reasoning

Learned Scanpaths Aid Blind Panoramic Video Quality Assessment

2024-03-30 · CVPR 2024 1 · Kanglong Fan, Wen Wen, Mu Li, Yifan Peng 외

Panoramic videos have the advantage of providing an immersive and interactive viewing experience. Nevertheless, their spherical nature gives rise to various and uncertain user viewing behaviors, which poses significant c…

Video Quality Assessment