paper-with-me

Papers

Unrolled Decomposed Unpaired Learning for Controllable Low-Light Video Enhancement

2024-08-22 · Lingyu Zhu, Wenhan Yang, Baoliang Chen, Hanwei Zhu, Zhangkai Ni, Qi Mao, Shiqi Wang

Obtaining pairs of low/normal-light videos, with motions, is more challenging than still images, which raises technical issues and poses the technical route of unpaired learning as a critical role. This paper makes endeavors in the direction of learning for low-light video enhancement without using paired ground truth. Compared to low-light image enhancement, enhancing low-light videos is more difficult due to the intertwined effects of noise, exposure, and contrast in the spatial domain, jointly with the need for temporal coherence. To address the above challenge, we propose the Unrolled Decomposed Unpaired Network (UDU-Net) for enhancing low-light videos by unrolling the optimization functions into a deep network to decompose the signal into spatial and temporal-related factors, which are updated iteratively. Firstly, we formulate low-light video enhancement as a Maximum A Posteriori estimation (MAP) problem with carefully designed spatial and temporal visual regularization. Then, via unrolling the problem, the optimization of the spatial and temporal constraints can be decomposed into different steps and updated in a stage-wise manner. From the spatial perspective, the designed Intra subnet leverages unpair prior information from expert photography retouched skills to adjust the statistical distribution. Additionally, we introduce a novel mechanism that integrates human perception feedback to guide network optimization, suppressing over/under-exposure conditions. Meanwhile, to address the issue from the temporal perspective, the designed Inter subnet fully exploits temporal cues in progressive optimization, which helps achieve improved temporal consistency in enhancement results. Consequently, the proposed method achieves superior performance to state-of-the-art methods in video illumination, noise suppression, and temporal consistency across outdoor and indoor scenes.

📄 PDF Abstract BibTeX arXiv:2408.12316

Code (1)

lingyzhu0101/udu 공식 구현 pytorch

Tasks

Image EnhancementLow-Light Image EnhancementVideo Enhancement

Similar Papers 제목 키워드 기반

UniTransfer: Video Concept Transfer via Progressive Spatial and Timestep Decomposition

2025-09-25 · Guojun Lei, Rong Zhang, Chi Wang, Tianhang Liu 외 arxiv

We propose a novel architecture UniTransfer, which introduces both spatial and diffusion timestep decomposition in a progressive paradigm, achieving precise and controllable video concept transfer. Specifically, in terms…

Representation Learning

LightCrafter: PBR-Conditioned Video Diffusion Refinement for Controllable and Consistent Relighting

2026-07-09 · Zixin Guo, Yehonathan Litman, Yifeng He, John Miller 외 arxiv

Video relighting requires balancing long-form temporal consistency with a physically grounded understanding of light transport, which depends on accurate estimation of intrinsic scene properties such as materials, geomet…

Inverse RenderingVideo Generation

MIMO: Controllable Character Video Synthesis with Spatial Decomposed Modeling

2024-09-24 · CVPR 2025 1 · Yifang Men, Yuan YAO, Miaomiao Cui, Liefeng Bo

Character video synthesis aims to produce realistic videos of animatable characters within lifelike scenes. As a fundamental problem in the computer vision and graphics community, 3D works typically require multi-view ca…

Wasserstein GANs for MR Imaging: from Paired to Unpaired Training

2019-10-15 · Ke Lei, Morteza Mardani, John M. Pauly, Shreyas S. Vasanawala

Lack of ground-truth MR images impedes the common supervised training of neural networks for image reconstruction. To cope with this challenge, this paper leverages unpaired adversarial training for reconstruction networ…

DiagnosticImage Reconstruction

NIR-assisted Video Enhancement via Unpaired 24-hour Data

2023-01-01 · ICCV 2023 1 · Muyao Niu, Zhihang Zhong, Yinqiang Zheng

Low-light video enhancement in the visible (VIS) range is important yet technically challenging, and it is likely to become more tractable by introducing near-infrared (NIR) information for assistance, which in turn …

Video Enhancement