paper-with-me

Papers

Predicting Visible Image Differences Under Varying Display Brightness and Viewing Distance

2019-06-01 · CVPR 2019 6 · Nanyang Ye, Krzysztof Wolski, Rafal K. Mantiuk

Numerous applications require a robust metric that can predict whether image differences are visible or not. However, the accuracy of existing white-box visibility metrics, such as HDR-VDP, is often not good enough. CNN-based black-box visibility metrics have proven to be more accurate, but they cannot account for differences in viewing conditions, such as display brightness and viewing distance. In this paper, we propose a CNN-based visibility metric, which maintains the accuracy of deep network solutions and accounts for viewing conditions. To achieve this, we extend the existing dataset of locally visible differences (LocVis) with a new set of measurements, collected considering aforementioned viewing conditions. Then, we develop a hybrid model that combines white-box processing stages for modeling the effects of luminance masking and contrast sensitivity, with a black-box deep neural network. We demonstrate that the novel hybrid model can handle the change of viewing conditions correctly and outperforms state-of-the-art metrics.

📄 PDF Abstract BibTeX

Code (2)

fanqiNO1/PyTorch-DPVM pytorch
ynyCL/DPVM tf

Similar Papers 제목 키워드 기반

Robust Perceptual Night Vision in Thermal Colorization

2020-03-04 · Feras Almasri, Olivier Debeir

Transforming a thermal infrared image into a robust perceptual colour Visible image is an ill-posed problem due to the differences in their spectral domains and in the objects' representations. Objects appear in one spec…

Colorization

XoFTR: Cross-modal Feature Matching Transformer

2024-04-15 · Önder Tuzcuoğlu, Aybora Köksal, Buğra Sofu, Sinan Kalkan 외

We introduce, XoFTR, a cross-modal cross-view method for local feature matching between thermal infrared (TIR) and visible images. Unlike visible images, TIR images are less susceptible to adverse lighting and weather co…

Image Augmentation

ShapeFormer: Shape Prior Visible-to-Amodal Transformer-based Amodal Instance Segmentation

2024-03-18 · Minh Tran, Winston Bounsavy, Khoa Vo, Anh Nguyen 외

Amodal Instance Segmentation (AIS) presents a challenging task as it involves predicting both visible and occluded parts of objects within images. Existing AIS methods rely on a bidirectional approach, encompassing both …

Amodal Instance SegmentationInstance SegmentationSemantic Segmentation

Learning Domain and Pose Invariance for Thermal-to-Visible Face Recognition

2022-11-17 · Cedric Nimpa Fondje, Shuowen Hu, Benjamin S. Riggan

Interest in thermal to visible face recognition has grown significantly over the last decade due to advancements in thermal infrared cameras and analytics beyond the visible spectrum. Despite large discrepancies between …

Face Recognition

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks

2025-08-11 · Yushen Xu, Xiaosong Li, Zhenyu Kuang, Xiaoqi Cheng 외 arxiv

The goal of multimodal image fusion is to integrate complementary information from infrared and visible images, generating multimodal fused images for downstream tasks. Existing downstream pre-training models are typical…

Semantic SegmentationObject Detection