Papers Video Forecasting
“Video Forecasting” 태그가 달린 논문 8편 · 필터 해제
Towards Efficient Real-Time Video Motion Transfer via Generative Time Series Modeling
We propose a deep learning framework designed to significantly optimize bandwidth for motion-transfer-enabled video applications, including video conferencing, virtual reality interactions, health monitoring systems, and…
Anomaly DetectionOptical Flow EstimationTime SeriesVideo ForecastingZAPBench: A Benchmark for Whole-Brain Activity Prediction in Zebrafish
Data-driven benchmarks have led to significant progress in key scientific modeling domains including weather and structural biology. Here, we introduce the Zebrafish Activity Prediction Benchmark (ZAPBench) to measure pr…
Activity PredictionMultivariate Time Series ForecastingSpatio-Temporal ForecastingTime Series Forecasting+1From Single to Multiple: Leveraging Multi-level Prediction Spaces for Video Forecasting
Despite video forecasting has been a widely explored topic in recent years, the mainstream of the existing work still limits their models with a single prediction space but completely neglects the way to leverage their m…
PredictionVideo ForecastingVideo PredictionTransformers in Vision: A Survey
Astounding results from Transformer models on natural language tasks have intrigued the vision community to study their application to computer vision problems. Among their salient benefits, Transformers enable modeling …
Action RecognitionActivity RecognitionColorizationimage-classification+14Sound2Sight: Generating Visual Dynamics from Sound and Context
Learning associations across modalities is critical for robust multimodal reasoning, especially when a modality may be missing during inference. In this paper, we study this problem in the context of audio-conditioned vi…
Multimodal ReasoningVideo ForecastingPredicting Future Instance Segmentation by Forecasting Convolutional Features
Anticipating future events is an important prerequisite towards intelligent behavior. Video forecasting has been studied as a proxy task towards this goal. Recent work has shown that to predict semantic segmentation of f…
Instance SegmentationOptical Flow EstimationSegmentationSemantic Segmentation+2Learning to Forecast Videos of Human Activity with Multi-granularity Models and Adaptive Rendering
We propose an approach for forecasting video of complex human activity involving multiple people. Direct pixel-level prediction is too simple to handle the appearance variability in complex activities. Hence, we develop …
DecoderVideo ForecastingThe Pose Knows: Video Forecasting by Generating Pose Futures
Current approaches in video forecasting attempt to generate videos directly in pixel space using Generative Adversarial Networks (GANs) or Variational Autoencoders (VAEs). However, since these approaches try to model all…
Human Pose ForecastingVideo ForecastingVideo Prediction