paper-with-me

홈 › Papers

Temporal Coherency based Criteria for Predicting Video Frames using Deep Multi-stage Generative Adversarial Networks

2017-12-01 · NeurIPS 2017 12 · Prateep Bhattacharjee, Sukhendu Das

Predicting the future from a sequence of video frames has been recently a sought after yet challenging task in the field of computer vision and machine learning. Although there have been efforts for tracking using motion trajectories and flow features, the complex problem of generating unseen frames has not been studied extensively. In this paper, we deal with this problem using convolutional models within a multi-stage Generative Adversarial Networks (GAN) framework. The proposed method uses two stages of GANs to generate a crisp and clear set of future frames. Although GANs have been used in the past for predicting the future, none of the works consider the relation between subsequent frames in the temporal dimension. Our main contribution lies in formulating two objective functions based on the Normalized Cross Correlation (NCC) and the Pairwise Contrastive Divergence (PCD) for solving this problem. This method, coupled with the traditional L1 loss, has been experimented with three real-world video datasets, viz. Sports-1M, UCF-101 and the KITTI. Performance analysis reveals superior results over the recent state-of-the-art methods.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Noisy-LSTM: Improving Temporal Awareness for Video Semantic Segmentation

2020-10-19 · Bowen Wang, Liangzhi Li, Yuta Nakashima, Ryo Kawasaki 외

Semantic video segmentation is a key challenge for various applications. This paper presents a new model named Noisy-LSTM, which is trainable in an end-to-end manner, with convolutional LSTMs (ConvLSTMs) to leverage the …

Semantic SegmentationVideo SegmentationVideo Semantic Segmentation

Temporally coherent video anonymization through GAN inpainting

2021-06-04 · Thangapavithraa Balaji, Patrick Blies, Georg Göri, Raphael Mitsch 외

This work tackles the problem of temporally coherent face anonymization in natural video streams.We propose JaGAN, a two-stage system starting with detecting and masking out faces with black image patches in all individu…

Face AnonymizationGenerative Adversarial NetworkPrivacy Preserving

LEO: Generative Latent Image Animator for Human Video Synthesis

2023-05-06 · Yaohui Wang, Xin Ma, Xinyuan Chen, Cunjian Chen 외

Spatio-temporal coherency is a major challenge in synthesizing high quality videos, particularly in synthesizing human videos that contain rich global and local deformations. To resolve this challenge, previous approache…

DisentanglementVideo Editing

Copy-and-Paste Networks for Deep Video Inpainting

2019-08-30 · ICCV 2019 10 · Sungho Lee, Seoung Wug Oh, DaeYeun Won, Seon Joo Kim

We present a novel deep learning based algorithm for video inpainting. Video inpainting is a process of completing corrupted or missing regions in videos. Video inpainting has additional challenges compared to image inpa…

Image InpaintingLane DetectionVideo Inpainting

Video Jigsaw: Unsupervised Learning of Spatiotemporal Context for Video Action Recognition

2018-08-22 · Unaiza Ahsan, Rishi Madhok, Irfan Essa

We propose a self-supervised learning method to jointly reason about spatial and temporal context for video recognition. Recent self-supervised approaches have used spatial context [9, 34] as well as temporal coherency […

Action RecognitionActivity RecognitionOptical Flow EstimationPosition+4