paper-with-me

홈 › Papers

Globally Consistent Video Depth and Pose Estimation with Efficient Test-Time Training

2022-08-04 · Yao-Chih Lee, Kuan-Wei Tseng, Guan-Sheng Chen, Chu-Song Chen

Dense depth and pose estimation is a vital prerequisite for various video applications. Traditional solutions suffer from the robustness of sparse feature tracking and insufficient camera baselines in videos. Therefore, recent methods utilize learning-based optical flow and depth prior to estimate dense depth. However, previous works require heavy computation time or yield sub-optimal depth results. We present GCVD, a globally consistent method for learning-based video structure from motion (SfM) in this paper. GCVD integrates a compact pose graph into the CNN-based optimization to achieve globally consistent estimation from an effective keyframe selection mechanism. It can improve the robustness of learning-based methods with flow-guided keyframes and well-established depth prior. Experimental results show that GCVD outperforms the state-of-the-art methods on both depth and pose estimation. Besides, the runtime experiments reveal that it provides strong efficiency in both short- and long-term videos with global consistency provided.

📄 PDF Abstract BibTeX arXiv:2208.02709

Code (1)

yaochih/gcvd-release 공식 구현 pytorch

Tasks

Optical Flow EstimationPose Estimation

Similar Papers 제목 키워드 기반

3D Consistency Optimization for Self-Supervised Monocular Video Depth Estimation

2026-06-14 · Yuanye Liu, Ke Zhang, Junzhe Jiang, Li Zhang 외 arxiv

Reliable monocular video depth estimation is crucial for downstream 3D reasoning and embodied AI in endoscopic navigation. However, existing self-supervised approaches typically treat video frames independently or rely o…

Multi-View 3D ReconstructionDepth Estimation

Unsupervised Scale-consistent Depth and Ego-motion Learning from Monocular Video

2019-08-28 · NeurIPS 2019 12 · Jia-Wang Bian, Zhichao Li, Naiyan Wang, Huangying Zhan 외

Recent work has shown that CNN-based depth and ego-motion estimators can be learned using unlabelled monocular videos. However, the performance is limited by unidentified moving objects that violate the underlying static…

Camera Pose EstimationDepth And Camera MotionDepth EstimationMonocular Depth Estimation+1

Align3R: Aligned Monocular Depth Estimation for Dynamic Videos

2024-12-04 · CVPR 2025 1 · Jiahao Lu, Tianyu Huang, Peng Li, Zhiyang Dou 외

Recent developments in monocular depth estimation methods enable high-quality depth estimation of single-view images but fail to estimate consistent video depth across different frames. Recent works address this problem …

Depth EstimationMonocular Depth Estimation

Novel View Synthesis of Dynamic Scenes with Globally Coherent Depths from a Monocular Camera

2020-04-02 · CVPR 2020 6 · Jae Shin Yoon, Kihwan Kim, Orazio Gallo, Hyun Soo Park 외

This paper presents a new method to synthesize an image from arbitrary views and times given a collection of images of a dynamic scene. A key challenge for the novel view synthesis arises from dynamic scene reconstructio…

Depth EstimationNovel View Synthesis

StereoDiff: Stereo-Diffusion Synergy for Video Depth Estimation

2025-06-25 · Haodong Li, Chen Wang, Jiahui Lei, Kostas Daniilidis 외

Recent video depth estimation methods achieve great performance by following the paradigm of image depth estimation, i.e., typically fine-tuning pre-trained video diffusion models with massive data. However, we argue tha…

Depth EstimationStereo Matching