Photo-Realistic Video Prediction on Natural Videos of Largely Changing Frames
Recent advances in deep learning have significantly improved performance of video prediction. However, state-of-the-art methods still suffer from blurriness and distortions in their future predictions, especially when there are large motions between frames. To address these issues, we propose a deep residual network with the hierarchical architecture where each layer makes a prediction of future state at different spatial resolution, and these predictions of different layers are merged via top-down connections to generate future frames. We trained our model with adversarial and perceptual loss functions, and evaluated it on a natural video dataset captured by car-mounted cameras. Our model quantitatively outperforms state-of-the-art baselines in future frame prediction on video sequences of both largely and slightly changing frames. Furthermore, our model generates future frames with finer details and textures that are perceptually more realistic than the baselines, especially under fast camera motions.
Code (0)
등록된 구현이 없습니다.
Tasks
PredictionVideo PredictionSimilar Papers 제목 키워드 기반
NPF-200: A Multi-Modal Eye Fixation Dataset and Method for Non-Photorealistic Videos
Non-photorealistic videos are in demand with the wave of the metaverse, but lack of sufficient research studies. This work aims to take a step forward to understand how humans perceive non-photorealistic videos with eye …
Saliency DetectionNLUT: Neural-based 3D Lookup Tables for Video Photorealistic Style Transfer
Video photorealistic style transfer is desired to generate videos with a similar photorealistic style to the style image while maintaining temporal consistency. However, existing methods obtain stylized video sequences b…
8kStyle TransferAutoVFX: Physically Realistic Video Editing from Natural Language Instructions
Modern visual effects (VFX) software has made it possible for skilled artists to create imagery of virtually anything. However, the creation process remains laborious, complex, and largely inaccessible to everyday users.…
Code GenerationVideo EditingPhotorealistic Style Transfer for Videos
Photorealistic style transfer is a technique which transfers colour from one reference domain to another domain by using deep learning and optimization techniques. Here, we present a technique which we use to transfer st…
Deep LearningStyle TransferAudio2Head: Audio-driven One-shot Talking-head Generation with Natural Head Motion
We propose an audio-driven talking-head method to generate photo-realistic talking-head videos from a single reference image. In this work, we tackle two key challenges: (i) producing natural head motions that match spee…
Image GenerationTalking Head Generation