paper-with-me

Papers

StereoPilot: Learning Unified and Efficient Stereo Conversion via Generative Priors

2025-12-18 · Guibao Shen, Yihua Du, Wenhang Ge, Jing He, Chirui Chang, Donghao Zhou, Zhen Yang, Luozhou Wang, Xin Tao, Ying-Cong Chen arxiv

The rapid growth of stereoscopic displays, including VR headsets and 3D cinemas, has led to increasing demand for high-quality stereo video content. However, producing 3D videos remains costly and complex, while automatic Monocular-to-Stereo conversion is hindered by the limitations of the multi-stage ``Depth-Warp-Inpaint'' (DWI) pipeline. This paradigm suffers from error propagation, depth ambiguity, and format inconsistency between parallel and converged stereo configurations. To address these challenges, we introduce UniStereo, the first large-scale unified dataset for stereo video conversion, covering both stereo formats to enable fair benchmarking and robust model training. Building upon this dataset, we propose StereoPilot, an efficient feed-forward model that directly synthesizes the target view without relying on explicit depth maps or iterative diffusion sampling. Equipped with a learnable domain switcher and a cycle consistency loss, StereoPilot adapts seamlessly to different stereo formats and achieves improved consistency. Extensive experiments demonstrate that StereoPilot significantly outperforms state-of-the-art methods in both visual fidelity and computational efficiency. Project page: https://hit-perfect.github.io/StereoPilot/.

📄 PDF Abstract BibTeX arXiv:2512.16915

Code (0)

등록된 구현이 없습니다.

Tasks

Computational Efficiency

Similar Papers 제목 키워드 기반

Beyond Feature Mapping GAP: Integrating Real HDRTV Priors for Superior SDRTV-to-HDRTV Conversion

2024-11-16 · Kepeng Xu, Li Xu, Gang He, Zhiqiang Zhang 외

The rise of HDR-WCG display devices has highlighted the need to convert SDRTV to HDRTV, as most video sources are still in SDR. Existing methods primarily focus on designing neural networks to learn a single-style mappin…

Generative Adversarial Network

Mono2Stereo: A Benchmark and Empirical Study for Stereo Conversion

2025-03-28 · CVPR 2025 1 · Songsong Yu, Yuxin Chen, Zhongang Qi, Zeke Xie 외

With the rapid proliferation of 3D devices and the shortage of 3D content, stereo conversion is attracting increasing attention. Recent works introduce pretrained Diffusion Models (DMs) into this task. However, due to th…

Lightweight Multiplane Images Network for Real-Time Stereoscopic Conversion from Planar Video

2024-12-04 · Shanding Diao, Yang Zhao, Yuan Chen, Zhao Zhang 외

With the rapid development of stereoscopic display technologies, especially glasses-free 3D screens, and virtual reality devices, stereoscopic conversion has become an important task to address the lack of high-quality s…

2k

SpatialMe: Stereo Video Conversion Using Depth-Warping and Blend-Inpainting

2024-12-16 · Jiale Zhang, Qianxi Jia, Yang Liu, Wei zhang 외

Stereo video conversion aims to transform monocular videos into immersive stereo format. Despite the advancements in novel view synthesis, it still remains two major challenges: i) difficulty of achieving high-fidelity a…

Novel View Synthesis

Elastic3D: Controllable Stereo Video Conversion with Guided Latent Decoding

2025-12-16 · Nando Metzger, Prune Truong, Goutam Bhat, Konrad Schindler 외 arxiv

The growing demand for immersive 3D content calls for automated monocular-to-stereo video conversion. We present Elastic3D, a controllable, direct end-to-end method for upgrading a conventional video to a binocular one. …

Depth Estimation