Learning Blind Video Temporal Consistency
Applying image processing algorithms independently to each frame of a video often leads to undesired inconsistent results over time. Developing temporally consistent video-based extensions, however, requires domain knowledge for individual tasks and is unable to generalize to other applications. In this paper, we present an efficient end-to-end approach based on deep recurrent network for enforcing temporal consistency in a video. Our method takes the original unprocessed and per-frame processed videos as inputs to produce a temporally consistent video. Consequently, our approach is agnostic to specific image processing algorithms applied on the original video. We train the proposed network by minimizing both short-term and long-term temporal losses as well as the perceptual loss to strike a balance between temporal stability and perceptual similarity with the processed frames. At test time, our model does not require computing optical flow and thus achieves real-time speed even for high-resolution videos. We show that our single model can handle multiple and unseen tasks, including but not limited to artistic style transfer, enhancement, colorization, image-to-image translation and intrinsic image decomposition. Extensive objective evaluation and subject study demonstrate that the proposed approach performs favorably against the state-of-the-art methods on various types of videos.
Code (1)
Tasks
ColorizationImage-to-Image TranslationIntrinsic Image DecompositionOptical Flow EstimationStyle TransferVideo Temporal ConsistencyMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Blind Video Temporal Consistency via Deep Video Prior
Applying image processing algorithms independently to each video frame often leads to temporal inconsistency in the resulting video. To address this issue, we present a novel and general approach for blind video temporal…
ColorizationImage DehazingImage EnhancementOptical Flow Estimation+2Frames2Residual: Spatiotemporal Decoupling for Self-Supervised Video Denoising
Self-supervised video denoising methods typically extend image-based frameworks into the temporal dimension, yet they often struggle to integrate inter-frame temporal consistency with intra-frame spatial specificity. Exi…
Video DenoisingTINQ: Temporal Inconsistency Guided Blind Video Quality Assessment
Blind video quality assessment (BVQA) has been actively researched for user-generated content (UGC) videos. Recently, super-resolution (SR) techniques have been widely applied in UGC. Therefore, an effective BVQA method …
Super-ResolutionVideo Quality AssessmentTemporal Kernel Consistency for Blind Video Super-Resolution
Deep learning-based blind super-resolution (SR) methods have recently achieved unprecedented performance in upscaling frames with unknown degradation. These models are able to accurately estimate the unknown downscaling …
Blind Super-ResolutionSuper-ResolutionVideo Super-ResolutionDeep Video Prior for Video Consistency and Propagation
Applying an image processing algorithm independently to each video frame often leads to temporal inconsistency in the resulting video. To address this issue, we present a novel and general approach for blind video tempor…
Optical Flow EstimationSemantic SegmentationVideo PropagationVideo Temporal Consistency