Predicting Visible Image Differences Under Varying Display Brightness and Viewing Distance
Numerous applications require a robust metric that can predict whether image differences are visible or not. However, the accuracy of existing white-box visibility metrics, such as HDR-VDP, is often not good enough. CNN-based black-box visibility metrics have proven to be more accurate, but they cannot account for differences in viewing conditions, such as display brightness and viewing distance. In this paper, we propose a CNN-based visibility metric, which maintains the accuracy of deep network solutions and accounts for viewing conditions. To achieve this, we extend the existing dataset of locally visible differences (LocVis) with a new set of measurements, collected considering aforementioned viewing conditions. Then, we develop a hybrid model that combines white-box processing stages for modeling the effects of luminance masking and contrast sensitivity, with a black-box deep neural network. We demonstrate that the novel hybrid model can handle the change of viewing conditions correctly and outperforms state-of-the-art metrics.
Code (2)
Similar Papers 제목 키워드 기반
Robust Perceptual Night Vision in Thermal Colorization
Transforming a thermal infrared image into a robust perceptual colour Visible image is an ill-posed problem due to the differences in their spectral domains and in the objects' representations. Objects appear in one spec…
ColorizationXoFTR: Cross-modal Feature Matching Transformer
We introduce, XoFTR, a cross-modal cross-view method for local feature matching between thermal infrared (TIR) and visible images. Unlike visible images, TIR images are less susceptible to adverse lighting and weather co…
Image AugmentationShapeFormer: Shape Prior Visible-to-Amodal Transformer-based Amodal Instance Segmentation
Amodal Instance Segmentation (AIS) presents a challenging task as it involves predicting both visible and occluded parts of objects within images. Existing AIS methods rely on a bidirectional approach, encompassing both …
Amodal Instance SegmentationInstance SegmentationSemantic SegmentationLearning Domain and Pose Invariance for Thermal-to-Visible Face Recognition
Interest in thermal to visible face recognition has grown significantly over the last decade due to advancements in thermal infrared cameras and analytics beyond the visible spectrum. Despite large discrepancies between …
Face RecognitionMambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks
The goal of multimodal image fusion is to integrate complementary information from infrared and visible images, generating multimodal fused images for downstream tasks. Existing downstream pre-training models are typical…
Semantic SegmentationObject Detection