paper-with-me

홈 › Papers

Online Overexposed Pixels Hallucination in Videos with Adaptive Reference Frame Selection

2023-08-29 · Yazhou Xing, Amrita Mazumdar, Anjul Patney, Chao Liu, Hongxu Yin, Qifeng Chen, Jan Kautz, Iuri Frosio

Low dynamic range (LDR) cameras cannot deal with wide dynamic range inputs, frequently leading to local overexposure issues. We present a learning-based system to reduce these artifacts without resorting to complex acquisition mechanisms like alternating exposures or costly processing that are typical of high dynamic range (HDR) imaging. We propose a transformer-based deep neural network (DNN) to infer the missing HDR details. In an ablation study, we show the importance of using a multiscale DNN and train it with the proper cost function to achieve state-of-the-art quality. To aid the reconstruction of the overexposed areas, our DNN takes a reference frame from the past as an additional input. This leverages the commonly occurring temporal instabilities of autoexposure to our advantage: since well-exposed details in the current frame may be overexposed in the future, we use reinforcement learning to train a reference frame selection DNN that decides whether to adopt the current frame as a future reference. Without resorting to alternating exposures, we obtain therefore a causal, HDR hallucination algorithm with potential application in common video acquisition settings. Our demo video can be found at https://drive.google.com/file/d/1-r12BKImLOYCLUoPzdebnMyNjJ4Rk360/view

📄 PDF Abstract BibTeX arXiv:2308.15462

Code (0)

등록된 구현이 없습니다.

Tasks

Hallucination

Similar Papers 제목 키워드 기반

HDR-cGAN: Single LDR to HDR Image Translation using Conditional GAN

2021-10-04 · Prarabdh Raipurkar, Rohil Pal, Shanmuganathan Raman

The prime goal of digital imaging techniques is to reproduce the realistic appearance of a scene. Low Dynamic Range (LDR) cameras are incapable of representing the wide dynamic range of the real-world scene. The captured…

HallucinationHDR ReconstructionTranslation

Online Adaptive Image Reconstruction (OnAIR) Using Dictionary Models

2018-09-06 · Brian E. Moore, Saiprasad Ravishankar, Raj Rao Nadakuditi, Jeffrey A. Fessler

Sparsity and low-rank models have been popular for reconstructing images and videos from limited or corrupted measurements. Dictionary or transform learning methods are useful in applications such as denoising, inpaintin…

DenoisingImage ReconstructionVideo Reconstruction

From Pixels to Tokens: Revisiting Object Hallucinations in Large Vision-Language Models

2024-10-09 · Yuying Shang, Xinyi Zeng, Yutao Zhu, Xiao Yang 외

Hallucinations in large vision-language models (LVLMs) are a significant challenge, i.e., generating objects that are not presented in the visual input, which impairs their reliability. Recent studies often attribute hal…

AttributeHallucination

SEASON: Mitigating Temporal Hallucination in Video Large Language Models via Self-Diagnostic Contrastive Decoding

2025-12-04 · Chang-Hsun Wu, Kai-Po Chang, Yu-Yang Sheng, Hung-Kai Chung 외 arxiv

Video Large Language Models (VideoLLMs) have shown remarkable progress in video understanding. However, these models still struggle to effectively perceive and exploit rich temporal information in videos when responding …

Dual Illumination Estimation for Robust Exposure Correction

2019-10-30 · Qing Zhang, Yongwei Nie, Wei-Shi Zheng

Exposure correction is one of the fundamental tasks in image processing and computational photography. While various methods have been proposed, they either fail to produce visually pleasing results, or only work well fo…

Exposure CorrectionMulti-Exposure Image Fusion