paper-with-me

홈 › Papers

Preserving Semantic and Temporal Consistency for Unpaired Video-to-Video Translation

2019-08-21 · Kwanyong Park, Sanghyun Woo, Dahun Kim, Donghyeon Cho, In So Kweon

In this paper, we investigate the problem of unpaired video-to-video translation. Given a video in the source domain, we aim to learn the conditional distribution of the corresponding video in the target domain, without seeing any pairs of corresponding videos. While significant progress has been made in the unpaired translation of images, directly applying these methods to an input video leads to low visual quality due to the additional time dimension. In particular, previous methods suffer from semantic inconsistency (i.e., semantic label flipping) and temporal flickering artifacts. To alleviate these issues, we propose a new framework that is composed of carefully-designed generators and discriminators, coupled with two core objective functions: 1) content preserving loss and 2) temporal consistency loss. Extensive qualitative and quantitative evaluations demonstrate the superior performance of the proposed method against previous approaches. We further apply our framework to a domain adaptation task and achieve favorable results.

📄 PDF Abstract BibTeX arXiv:1908.07683

Code (0)

등록된 구현이 없습니다.

Tasks

Domain AdaptationTranslation

Similar Papers 제목 키워드 기반

Learning Temporally and Semantically Consistent Unpaired Video-to-video Translation Through Pseudo-Supervision From Synthetic Optical Flow

2022-01-15 · Kaihong Wang, Kumar Akash, Teruhisa Misu

Unpaired video-to-video translation aims to translate videos between a source and a target domain without the need of paired training data, making it more feasible for real applications. Unfortunately, the translated vid…

Motion EstimationOptical Flow EstimationTranslation

Long-Term Temporally Consistent Unpaired Video Translation from Simulated Surgical 3D Data

2021-03-31 · ICCV 2021 10 · Dominik Rivoir, Micha Pfeiffer, Reuben Docea, Fiona Kolbinger 외

Research in unpaired video translation has mainly focused on short-term temporal consistency by conditioning on neighboring frames. However for transfer from simulated to photorealistic sequences, available information o…

Neural RenderingTranslation

Beyond Consistency: Preserving Temporal Structure in Zero-Shot Video Editing

2026-06-07 · Deyin Liu, Yisheng Ding, Zhe Jin, Xiatian Zhu 외 arxiv

Existing zero-shot video editing methods rely on pre-trained diffusion models, successfully achieving spatial control and basic temporal consistency but fundamentally fail to preserve the video's original temporal struct…

Computational Efficiency

NOVA: Sparse Control, Dense Synthesis for Pair-Free Video Editing

2026-03-03 · Tianlin Pan, Jiayi Dai, Chenpu Yuan, Zhengyao Lv 외 arxiv

Recent video editing models have achieved impressive results, but most still require large-scale paired datasets. Collecting such naturally aligned pairs at scale remains highly challenging and constitutes a critical bot…

Image Editing

I2V-GAN: Unpaired Infrared-to-Visible Video Translation

2021-08-02 · Shuang Li, Bingfeng Han, Zhenjie Yu, Chi Harold Liu 외

Human vision is often adversely affected by complex environmental factors, especially in night vision scenarios. Thus, infrared cameras are often leveraged to help enhance the visual effects via detecting infrared radiat…

object-detectionObject DetectionTranslation