paper-with-me

홈 › Papers

ImmersePro: End-to-End Stereo Video Synthesis Via Implicit Disparity Learning

2024-09-30 · Jian Shi, Zhenyu Li, Peter Wonka

We introduce \textit{ImmersePro}, an innovative framework specifically designed to transform single-view videos into stereo videos. This framework utilizes a novel dual-branch architecture comprising a disparity branch and a context branch on video data by leveraging spatial-temporal attention mechanisms. \textit{ImmersePro} employs implicit disparity guidance, enabling the generation of stereo pairs from video sequences without the need for explicit disparity maps, thus reducing potential errors associated with disparity estimation models. In addition to the technical advancements, we introduce the YouTube-SBS dataset, a comprehensive collection of 423 stereo videos sourced from YouTube. This dataset is unprecedented in its scale, featuring over 7 million stereo pairs, and is designed to facilitate training and benchmarking of stereo video generation models. Our experiments demonstrate the effectiveness of \textit{ImmersePro} in producing high-quality stereo videos, offering significant improvements over existing methods. Compared to the best competitor stereo-from-mono we quantitatively improve the results by 11.76\% (L1), 6.39\% (SSIM), and 5.10\% (PSNR).

📄 PDF Abstract BibTeX arXiv:2410.00262

Code (1)

shijianjian/ImmersePro 공식 구현 pytorch

Tasks

BenchmarkingDisparity EstimationSSIMVideo Generation

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음

Similar Papers 제목 키워드 기반

Neural Stereo Video Compression with Hybrid Disparity Compensation

2025-04-29 · Shiyin Jiang, Zhenghao Chen, Minghao Han, Xingyu Zhou 외

Disparity compensation represents the primary strategy in stereo video compression (SVC) for exploiting cross-view redundancy. These mechanisms can be broadly categorized into two types: one that employs explicit horizon…

Autonomous DrivingVideo Compression

Stereo World Model: Camera-Guided Stereo Video Generation

2026-03-18 · Yang-Tian Sun, Zehuan Huang, Yifan Niu, Lin Ma 외 arxiv

We present StereoWorld, a camera-conditioned stereo world model that jointly learns appearance and binocular geometry for end-to-end stereo video generation.Unlike monocular RGB or RGBD approaches, StereoWorld operates e…

Depth EstimationVideo Generation

Eye2Eye: A Simple Approach for Monocular-to-Stereo Video Synthesis

2025-04-30 · Michal Geyer, Omer Tov, Linyi Jin, Richard Tucker 외

The rising popularity of immersive visual experiences has increased interest in stereoscopic 3D video generation. Despite significant advances in video synthesis, creating 3D videos remains challenging due to the relativ…

Disparity EstimationTransparent objectsVideo Generation

Stereo-Knowledge Distillation from dpMV to Dual Pixels for Light Field Video Reconstruction

2024-05-20 · Aryan Garg, Raghav Mallampali, Akshat Joshi, Shrisudhan Govindarajan 외

Dual pixels contain disparity cues arising from the defocus blur. This disparity information is useful for many vision tasks ranging from autonomous driving to 3D creative realism. However, directly estimating disparity …

Autonomous DrivingKnowledge DistillationVideo Reconstruction

SpatialMe: Stereo Video Conversion Using Depth-Warping and Blend-Inpainting

2024-12-16 · Jiale Zhang, Qianxi Jia, Yang Liu, Wei zhang 외

Stereo video conversion aims to transform monocular videos into immersive stereo format. Despite the advancements in novel view synthesis, it still remains two major challenges: i) difficulty of achieving high-fidelity a…

Novel View Synthesis