paper-with-me

Papers

Learning Generalized Spatial-Temporal Deep Feature Representation for No-Reference Video Quality Assessment

2020-12-27 · Baoliang Chen, Lingyu Zhu, Guo Li, Hongfei Fan, Shiqi Wang

In this work, we propose a no-reference video quality assessment method, aiming to achieve high-generalization capability in cross-content, -resolution and -frame rate quality prediction. In particular, we evaluate the quality of a video by learning effective feature representations in spatial-temporal domain. In the spatial domain, to tackle the resolution and content variations, we impose the Gaussian distribution constraints on the quality features. The unified distribution can significantly reduce the domain gap between different video samples, resulting in a more generalized quality feature representation. Along the temporal dimension, inspired by the mechanism of visual perception, we propose a pyramid temporal aggregation module by involving the short-term and long-term memory to aggregate the frame-level quality. Experiments show that our method outperforms the state-of-the-art methods on cross-dataset settings, and achieves comparable performance on intra-dataset configurations, demonstrating the high-generalization capability of the proposed method.

📄 PDF Abstract BibTeX arXiv:2012.13936

Code (1)

Baoliang93/GSTVQA 공식 구현 pytorch

Tasks

Video Quality Assessment

Similar Papers 제목 키워드 기반

Pyramid Spatial-Temporal Aggregation for Video-Based Person Re-Identification

2021-01-01 · ICCV 2021 10 · Yingquan Wang, Pingping Zhang, Shang Gao, Xia Geng 외

Video-based person re-identification aims to associate the video clips of the same person across multiple non-overlapping cameras. Spatial-temporal representations can provide richer and complementary information bet…

Person Re-IdentificationVideo-Based Person Re-Identification

Video Quality Assessment Based on Swin TransformerV2 and Coarse to Fine Strategy

2024-01-16 · Zihao Yu, Fengbin Guan, Yiting Lu, Xin Li 외

The objective of non-reference video quality assessment is to evaluate the quality of distorted video without access to reference high-definition references. In this study, we introduce an enhanced spatial perception mod…

Image Quality AssessmentVideo Quality AssessmentVisual Question Answering (VQA)

ST-GREED: Space-Time Generalized Entropic Differences for Frame Rate Dependent Video Quality Prediction

2020-10-26 · Pavan C. Madhusudana, Neil Birkbeck, Yilin Wang, Balu Adsumilli 외

We consider the problem of conducting frame rate dependent video quality assessment (VQA) on videos of diverse frame rates, including high frame rate (HFR) videos. More generally, we study how perceptual quality is affec…

Video Quality AssessmentVisual Question Answering (VQA)

Blockwise Temporal-Spatial Pathway Network

2022-08-05 · SeulGi Hong, Min-Kook Choi

Algorithms for video action recognition should consider not only spatial information but also temporal relations, which remains challenging. We propose a 3D-CNN-based action recognition model, called the blockwise tempor…

Action RecognitionTemporal Action Localization

TSception: Capturing Temporal Dynamics and Spatial Asymmetry from EEG for Emotion Recognition

2021-04-07 · Yi Ding, Neethu Robinson, Su Zhang, Qiuhao Zeng 외

The high temporal resolution and the asymmetric spatial activations are essential attributes of electroencephalogram (EEG) underlying emotional processes in the brain. To learn the temporal dynamics and spatial asymmetry…

EEGEmotion Recognition