paper-with-me

Papers

Better Guider Predicts Future Better: Difference Guided Generative Adversarial Networks

2019-01-07 · Guohao Ying, Yingtian Zou, Lin Wan, Yiming Hu, Jiashi Feng

Predicting the future is a fantasy but practicality work. It is the key component to intelligent agents, such as self-driving vehicles, medical monitoring devices and robotics. In this work, we consider generating unseen future frames from previous obeservations, which is notoriously hard due to the uncertainty in frame dynamics. While recent works based on generative adversarial networks (GANs) made remarkable progress, there is still an obstacle for making accurate and realistic predictions. In this paper, we propose a novel GAN based on inter-frame difference to circumvent the difficulties. More specifically, our model is a multi-stage generative network, which is named the Difference Guided Generative Adversarial Netwok (DGGAN). The DGGAN learns to explicitly enforce future-frame predictions that is guided by synthetic inter-frame difference. Given a sequence of frames, DGGAN first uses dual paths to generate meta information. One path, called Coarse Frame Generator, predicts the coarse details about future frames, and the other path, called Difference Guide Generator, generates the difference image which include complementary fine details. Then our coarse details will then be refined via guidance of difference image under the support of GANs. With this model and novel architecture, we achieve state-of-the-art performance for future video prediction on UCF-101, KITTI.

📄 PDF Abstract BibTeX arXiv:1901.01649

Code (0)

등록된 구현이 없습니다.

Tasks

Video Prediction

Methods 이 논문이 사용한 방법론

Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…
Dogecoin Customer Service Number +1-833-534-1729 설명 없음

Similar Papers 제목 키워드 기반

Probabilistic Human Intent Prediction for Mobile Manipulation: An Evaluation with Human-Inspired Constraints

2025-07-14 · Cesar Alan Contreras, Manolis Chiou, Alireza Rastegarpanah, Michal Szulik 외 arxiv

Accurate inference of human intent enables human-robot collaboration without constraining human control or causing conflicts between humans and robots. We present GUIDER (Global User Intent Dual-phase Estimation for Robo…

VR IQA NET: Deep Virtual Reality Image Quality Assessment using Adversarial Learning

2018-04-11 · Heoun-taek Lim, Hak Gu Kim, Yong Man Ro

In this paper, we propose a novel virtual reality image quality assessment (VR IQA) with adversarial learning for omnidirectional images. To take into account the characteristics of the omnidirectional image, we devise d…

Image Quality AssessmentPosition

LMVP: Video Predictor with Leaked Motion Information

2019-06-24 · Dong Wang, Yitong Li, Wei Cao, Liqun Chen 외

We propose a Leaked Motion Video Predictor (LMVP) to predict future frames by capturing the spatial and temporal dependencies from given inputs. The motion is modeled by a newly proposed component, motion guider, which p…

Utilizing Vision-Language Models as Action Models for Intent Recognition and Assistance

2025-08-14 · Cesar Alan Contreras, Manolis Chiou, Alireza Rastegarpanah, Michal Szulik 외 arxiv

Human-robot collaboration requires robots to quickly infer user intent, provide transparent reasoning, and assist users in achieving their goals. Our recent work introduced GUIDER, our framework for inferring navigation …

Instance SegmentationIntent RecognitionObject Detection

Pose Transferrable Person Re-Identification

2018-06-01 · CVPR 2018 6 · Jinxian Liu, Bingbing Ni, Yichao Yan, Peng Zhou 외

Person re-identification (ReID) is an important task in the field of intelligent security. A key challenge is how to capture human pose variations, while existing benchmarks (i.e., Market1501, DukeMTMC-reID, CUHK03, etc.…

Person Re-IdentificationTriplet