paper-with-me

Papers

Video Reconstruction using Diffusion-based Image-to-Video Generation with Trajectory Guidance

2026-05-14 · Stelio Bompai, Ioannis Kontopoulos, Giannis Spiliopoulos, Dimitris Zissis, Konstantinos Tserpes arxiv

This paper addresses the problem of reconstructing missing or dropped frames in top-down drone video of autonomous surface vehicles performing structured maritime manoeuvres. We propose a pipeline that converts raw GPS telemetry and a single reference frame into a trajectory-guided video sequence using a pre-trained image-to-video diffusion model, requiring no domain-specific fine-tuning. GPS coordinates from onboard telemetry logs are projected into image space via an equirectangular mapping, producing per-vessel motion cues that condition the SG-I2V diffusion model. The generated frames are evaluated against ground-truth video using perceptual, temporal and trajectory-based metrics, and benchmarked against optical flow extrapolation and RIFE interpolation baselines. SG-I2V produces the most naturally appearing frames among all methods (BRISQUE 25.52, closest to ground-truth 23.64), the most realistic motion magnitude (temporal smoothness 1.14 vs. ground truth 1.42), and the strongest GPS trajectory adherence (9.31px vs. 28.70px for ground-truth, the latter reflecting approximate temporal alignment between footage and GPS logs rather than generation error), demonstrating that trajectory-guided diffusion synthesis is a viable approach to maritime video reconstruction under challenging low-texture, small-object conditions.

📄 PDF Abstract BibTeX arXiv:2605.16420

Code (0)

등록된 구현이 없습니다.

Tasks

Video ReconstructionVideo Generation

Similar Papers 제목 키워드 기반

Hi3D: Pursuing High-Resolution Image-to-3D Generation with Video Diffusion Models

2024-09-11 · Haibo Yang, Yang Chen, Yingwei Pan, Ting Yao 외

Despite having tremendous progress in image-to-3D generation, existing methods still struggle to produce multi-view consistent images with high-resolution textures in detail, especially in the paradigm of 2D diffusion th…

3D Generation3D ReconstructionImage GenerationImage to 3D+2

HoloTime: Taming Video Diffusion Models for Panoramic 4D Scene Generation

2025-04-30 · Haiyang Zhou, Wangbo Yu, Jiawen Guan, Xinhua Cheng 외

The rapid advancement of diffusion models holds the promise of revolutionizing the application of VR and AR technologies, which typically require scene-level 4D assets for user experience. Nonetheless, existing diffusion…

Depth EstimationScene GenerationVideo Generation

LaMD: Latent Motion Diffusion for Image-Conditional Video Generation

2023-04-23 · Yaosi Hu, Zhenzhong Chen, Chong Luo

The video generation field has witnessed rapid improvements with the introduction of recent diffusion models. While these models have successfully enhanced appearance quality, they still face challenges in generating coh…

Motion GenerationVideo GenerationVideo Reconstruction

Wonderland: Navigating 3D Scenes from a Single Image

2024-12-16 · CVPR 2025 1 · Hanwen Liang, Junli Cao, Vidit Goel, Guocheng Qian 외

This paper addresses a challenging question: How can we efficiently create high-quality, wide-scope 3D scenes from a single arbitrary image? Existing methods face several constraints, such as requiring multi-view data, t…

3D ReconstructionScene Generation

SV3D: Novel Multi-view Synthesis and 3D Generation from a Single Image using Latent Video Diffusion

2024-03-18 · Vikram Voleti, Chun-Han Yao, Mark Boss, Adam Letts 외

We present Stable Video 3D (SV3D) -- a latent video diffusion model for high-resolution, image-to-multi-view generation of orbital videos around a 3D object. Recent work on 3D generation propose techniques to adapt 2D ge…

3D Generation3D ReconstructionImage to 3DNovel View Synthesis