paper-with-me

Papers

ColorVideoVDP: A visual difference predictor for image, video and display distortions

2024-01-21 · Rafal K. Mantiuk, Param Hanji, Maliha Ashraf, Yuta Asano, ALEXANDRE CHAPIRO

ColorVideoVDP is a video and image quality metric that models spatial and temporal aspects of vision, for both luminance and color. The metric is built on novel psychophysical models of chromatic spatiotemporal contrast sensitivity and cross-channel contrast masking. It accounts for the viewing conditions, geometric, and photometric characteristics of the display. It was trained to predict common video streaming distortions (e.g. video compression, rescaling, and transmission errors), and also 8 new distortion types related to AR/VR displays (e.g. light source and waveguide non-uniformities). To address the latter application, we collected our novel XR-Display-Artifact-Video quality dataset (XR-DAVID), comprised of 336 distorted videos. Extensive testing on XR-DAVID, as well as several datasets from the literature, indicate a significant gain in prediction performance compared to existing metrics. ColorVideoVDP opens the doors to many novel applications which require the joint automated spatiotemporal assessment of luminance and color distortions, including video streaming, display specification and design, visual comparison of results, and perceptually-guided quality optimization.

📄 PDF Abstract BibTeX arXiv:2401.11485

Code (1)

gfxdisp/colorvideovdp 공식 구현 pytorch

Tasks

Video Compression

Similar Papers 제목 키워드 기반

HDR-VDP-3: A multi-metric for predicting image differences, quality and contrast distortions in high dynamic range and regular content

2023-04-26 · Rafal K. Mantiuk, Dounia Hammou, Param Hanji

High-Dynamic-Range Visual-Difference-Predictor version 3, or HDR-VDP-3, is a visual metric that can fulfill several tasks, such as full-reference image/video quality assessment, prediction of visual differences between a…

PositionPredictionVideo Quality Assessment

Towards Visual-Prompt Temporal Answering Grounding in Medical Instructional Video

2022-03-13 · Bin Li, Yixuan Weng, Bin Sun, Shutao Li

The temporal answering grounding in the video (TAGV) is a new task naturally derived from temporal sentence grounding in the video (TSGV). Given an untrimmed video and a text question, this task aims at locating the matc…

Language ModellingQuestion AnsweringSentenceTemporal Sentence Grounding

FovVideoVDP: A visible difference predictor for wide field-of-view video

2021-04-01 · ACM Transactions on Graphics 2021 4 · Rafał K. Mantiuk, Gyorgy Denes, ALEXANDRE CHAPIRO, Anton Kaplanyan 외

FovVideoVDP is a video difference metric that models the spatial, temporal, and peripheral aspects of perception. While many other metrics are available, our work provides the first practical treatment of these three cen…

Sensitivity

CameraVDP: Perceptual Display Assessment with Uncertainty Estimation via Camera and Visual Difference Prediction

2025-09-10 · Yancheng Cai, Robert Wanat, Rafal Mantiuk arxiv

Accurate measurement of images produced by electronic displays is critical for the evaluation of both traditional and computational displays. Traditional display measurement methods based on sparse radiometric sampling a…

Learning Depth from Monocular Videos using Direct Methods

2017-12-01 · CVPR 2018 6 · Chaoyang Wang, Jose Miguel Buenaposada, Rui Zhu, Simon Lucey

The ability to predict depth from a single image - using recent advances in CNNs - is of increasing interest to the vision community. Unsupervised strategies to learning are particularly appealing as they can utilize muc…

Depth And Camera MotionVisual Odometry