paper-with-me

홈 › Papers

Visual Dynamics: Stochastic Future Generation via Layered Cross Convolutional Networks

2018-07-24 · Tianfan Xue, Jiajun Wu, Katherine L. Bouman, William T. Freeman

We study the problem of synthesizing a number of likely future frames from a single input image. In contrast to traditional methods that have tackled this problem in a deterministic or non-parametric way, we propose to model future frames in a probabilistic manner. Our probabilistic model makes it possible for us to sample and synthesize many possible future frames from a single input image. To synthesize realistic movement of objects, we propose a novel network structure, namely a Cross Convolutional Network; this network encodes image and motion information as feature maps and convolutional kernels, respectively. In experiments, our model performs well on synthetic data, such as 2D shapes and animated game sprites, and on real-world video frames. We present analyses of the learned network representations, showing it is implicitly learning a compact encoding of object appearance and motion. We also demonstrate a few of its applications, including visual analogy-making and video extrapolation.

📄 PDF Abstract BibTeX arXiv:1807.09245

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

The Strategic LQG System: A Dynamic Stochastic VCG Framework for Optimal Coordination

2019-06-11

The classic Vickrey-Clarke-Groves (VCG) mechanism ensures incentive compatibility, i.e., that truth-telling of all agents is a dominant strategy, for a static one-shot game. However, in a dynamic environment that unfolds…

Text2Layer: Layered Image Generation using Latent Diffusion Model

2023-07-19 · Xinyang Zhang, Wentian Zhao, Xin Lu, Jeff Chien

Layer compositing is one of the most popular image editing workflows among both amateurs and professionals. Motivated by the success of diffusion models, we explore layer compositing from a layered image generation persp…

Image GenerationImage SegmentationmodelSemantic Segmentation

Vera: A Layered Diffusion Model for Content-Preserving Video Editing

2026-06-22 · Hongkai Zheng, Ta-Ying Cheng, Benjamin Klein, Yisong Yue 외 arxiv

Video diffusion models have enabled remarkable progress in video generation and editing. However, content preservation remains a core challenge: existing methods regenerate every pixel and often alter elements that shoul…

Video Generation

A Unified and Controllable Framework for Layered Image Generation with Visual Effects

2026-01-21 · Jinrui Yang, Qing Liu, Yijun Li, Mengwei Ren 외 arxiv

Recent image generation models produce impressive composites, but often fail to preserve the identity of user-provided content when editing specific elements: the surrounding scene may shift, and even the edited object's…

Image Generation

Visual Path Prediction in Complex Scenes With Crowded Moving Objects

2016-06-01 · CVPR 2016 6 · YoungJoon Yoo, Kimin Yun, Sangdoo Yun, JongHee Hong 외

This paper proposes a novel path prediction algorithm for progressing one step further than the existing works focusing on single target path prediction. In this paper, we consider moving dynamics of co-occurring objects…

Prediction