paper-with-me

홈 › Papers

SpatioTemporal Feature Integration and Model Fusion for Full Reference Video Quality Assessment

2018-04-13

Perceptual video quality assessment models are either frame-based or video-based, i.e., they apply spatiotemporal filtering or motion estimation to capture temporal video distortions. Despite their good performance on video quality databases, video-based approaches are time-consuming and harder to efficiently deploy. To balance between high performance and computational efficiency, Netflix developed the Video Multi-method Assessment Fusion (VMAF) framework, which integrates multiple quality-aware features to predict video quality. Nevertheless, this fusion framework does not fully exploit temporal video quality measurements which are relevant to temporal video distortions. To this end, we propose two improvements to the VMAF framework: SpatioTemporal VMAF and Ensemble VMAF. Both algorithms exploit efficient temporal video features which are fed into a single or multiple regression models. To train our models, we designed a large subjective database and evaluated the proposed models against state-of-the-art approaches. The compared algorithms will be made available as part of the open source package in https://github.com/Netflix/vmaf.

📄 PDF Abstract BibTeX arXiv:1804.04813

Code (1)

Netflix/vmaf 공식 구현

Tasks

Computational EfficiencyMotion EstimationVideo Quality Assessment

Similar Papers 제목 키워드 기반

Bridge Frame and Event: Common Spatiotemporal Fusion for High-Dynamic Scene Optical Flow

2025-03-10 · CVPR 2025 1 · Hanyu Zhou, Haonan Wang, Haoyue Liu, Yuxing Duan 외

High-dynamic scene optical flow is a challenging task, which suffers spatial blur and temporal discontinuous motion due to large displacement in frame imaging, thus deteriorating the spatiotemporal feature of optical flo…

Optical Flow Estimation

CardioDiT: Latent Diffusion Transformers for 4D Cardiac MRI Synthesis

2026-03-26 · Marvin Seyfarth, Sarah Kaye Müller, Arman Ghanaat, Isabelle Ayx 외 arxiv

Latent diffusion models (LDMs) have recently achieved strong performance in 3D medical image synthesis. However, modalities like cine cardiac MRI (CMR), representing a temporally synchronized 3D volume across the cardiac…

C3DVQA: Full-Reference Video Quality Assessment with 3D Convolutional Neural Network

2019-10-30 · Munan Xu, Junming Chen, Haiqiang Wang, Shan Liu 외

Traditional video quality assessment (VQA) methods evaluate localized picture quality and video score is predicted by temporally aggregating frame scores. However, video quality exhibits different characteristics from st…

Video Quality AssessmentVisual Question Answering (VQA)

A Decade of Deep Learning for Remote Sensing Spatiotemporal Fusion: Advances, Challenges, and Opportunities

2025-04-01 · Enzhe Sun, Yongchuan Cui, Peng Liu, Jining Yan

Hardware limitations and satellite launch costs make direct acquisition of high temporal-spatial resolution remote sensing imagery challenging. Remote sensing spatiotemporal fusion (STF) technology addresses this problem…

Spatiotemporal Pyramid Network for Video Action Recognition

2019-03-04 · CVPR 2017 7 · Yunbo Wang, Mingsheng Long, Jian-Min Wang, Philip S. Yu

Two-stream convolutional networks have shown strong performance in video action recognition tasks. The key idea is to learn spatiotemporal features by fusing convolutional networks spatially and temporally. However, it r…

Action RecognitionTemporal Action Localization