paper-with-me

홈 › Papers

Geometry-Based Next Frame Prediction from Monocular Video

2016-09-20 · Reza Mahjourian, Martin Wicke, Anelia Angelova

We consider the problem of next frame prediction from video input. A recurrent convolutional neural network is trained to predict depth from monocular video input, which, along with the current video image and the camera trajectory, can then be used to compute the next frame. Unlike prior next-frame prediction approaches, we take advantage of the scene geometry and use the predicted depth for generating the next frame prediction. Our approach can produce rich next frame predictions which include depth information attached to each pixel. Another novel aspect of our approach is that it predicts depth from a sequence of images (e.g. in a video), rather than from a single still image. We evaluate the proposed approach on the KITTI dataset, a standard dataset for benchmarking tasks relevant to autonomous driving. The proposed method produces results which are visually and numerically superior to existing methods that directly predict the next frame. We show that the accuracy of depth prediction improves as more prior frames are considered.

📄 PDF Abstract BibTeX arXiv:1609.06377

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous DrivingBenchmarkingDepth EstimationDepth PredictionPrediction

Similar Papers 제목 키워드 기반

Video Generative Models as Geometry Learner

2026-08-28 · Haosen Yang, Jifei Song, Zhensong Zhang, Xiatian Zhu 외 arxiv

Recent generative approaches to geometry estimation adapt pretrained image diffusion models and treat the task as image-conditioned generation. Leveraging off-the-shelf image diffusion models, they either (i) train task-…

Pseudo RGB-D for Self-Improving Monocular SLAM and Depth Prediction

2020-04-22 · ECCV 2020 8 · Lokender Tiwari, Pan Ji, Quoc-Huy Tran, Bingbing Zhuang 외

Classical monocular Simultaneous Localization And Mapping (SLAM) and the recently emerging convolutional neural networks (CNNs) for monocular depth prediction represent two largely disjoint approaches towards building a …

Depth EstimationDepth PredictionSimultaneous Localization and Mapping

MonoPP: Metric-Scaled Self-Supervised Monocular Depth Estimation by Planar-Parallax Geometry in Automotive Applications

2024-11-29 · Gasser Elazab, Torben Gräber, Michael Unterreiner, Olaf Hellwich

Self-supervised monocular depth estimation (MDE) has gained popularity for obtaining depth predictions directly from videos. However, these methods often produce scale invariant results, unless additional training signal…

Depth EstimationDepth PredictionMonocular Depth EstimationPosition

Predicting 3D representations for Dynamic Scenes

2025-01-28 · Di Qi, Tong Yang, Beining Wang, Xiangyu Zhang 외

We present a novel framework for dynamic radiance field prediction given monocular video streams. Unlike previous methods that primarily focus on predicting future frames, our method goes a step further by generating exp…

Learning Multi-frame and Monocular Prior for Estimating Geometry in Dynamic Scenes

2025-05-03 · Seong Hyeon Park, Jinwoo Shin

In monocular videos that capture dynamic scenes, estimating the 3D geometry of video contents has been a fundamental challenge in computer vision. Specifically, the task is significantly challenged by the object motion, …

3D geometry