paper-with-me

홈 › Papers

CoProU-VO: Combining Projected Uncertainty for End-to-End Unsupervised Monocular Visual Odometry

2025-08-01 · Jingchao Xie, Oussema Dhaouadi, Weirong Chen, Johannes Meier, Jacques Kaiser, Daniel Cremers arxiv

Visual Odometry (VO) is fundamental to autonomous navigation, robotics, and augmented reality, with unsupervised approaches eliminating the need for expensive ground-truth labels. However, these methods struggle when dynamic objects violate the static scene assumption, leading to erroneous pose estimations. We tackle this problem by uncertainty modeling, which is a commonly used technique that creates robust masks to filter out dynamic objects and occlusions without requiring explicit motion segmentation. Traditional uncertainty modeling considers only single-frame information, overlooking the uncertainties across consecutive frames. Our key insight is that uncertainty must be propagated and combined across temporal frames to effectively identify unreliable regions, particularly in dynamic scenes. To address this challenge, we introduce Combined Projected Uncertainty VO (CoProU-VO), a novel end-to-end approach that combines target frame uncertainty with projected reference frame uncertainty using a principled probabilistic formulation. Built upon vision transformer backbones, our model simultaneously learns depth, uncertainty estimation, and camera poses. Consequently, experiments on the KITTI and nuScenes datasets demonstrate significant improvements over previous unsupervised monocular end-to-end two-frame-based methods and exhibit strong performance in challenging highway scenes where other approaches often fail. Additionally, comprehensive ablation studies validate the effectiveness of cross-frame uncertainty propagation.

📄 PDF Abstract BibTeX arXiv:2508.00568

Code (0)

등록된 구현이 없습니다.

Tasks

Motion SegmentationVisual Odometry

Similar Papers 제목 키워드 기반

3D Distillation: Improving Self-Supervised Monocular Depth Estimation on Reflective Surfaces

2023-01-01 · ICCV 2023 1 · Xuepeng Shi, Georgi Dikov, Gerhard Reitmayr, Tae-Kyun Kim 외

Self-supervised monocular depth estimation (SSMDE) aims at predicting the dense depth maps of monocular images, by learning to minimize a photometric loss using spatially neighboring image pairs during training. Whil…

Depth EstimationMonocular Depth Estimation

MonoInstance: Enhancing Monocular Priors via Multi-view Instance Alignment for Neural Rendering and Reconstruction

2025-03-24 · CVPR 2025 1 · Wenyuan Zhang, Yixiao Yang, Han Huang, Liang Han 외

Monocular depth priors have been widely adopted by neural rendering in multi-view based tasks such as 3D reconstruction and novel view synthesis. However, due to the inconsistent prediction on each view, how to more effe…

3D ReconstructionNeural RenderingNovel View Synthesis

MonoProb: Self-Supervised Monocular Depth Estimation with Interpretable Uncertainty

2023-11-10 · Remi Marsal Florian Chabot, Angelique Loesch, William Grolleau, Hichem Sahbi

Self-supervised monocular depth estimation methods aim to be used in critical applications such as autonomous vehicles for environment analysis. To circumvent the potential imperfections of these approaches, a quantifica…

Autonomous VehiclesDecision MakingDepth EstimationDepth Prediction+2

GUPNet++: Geometry Uncertainty Propagation Network for Monocular 3D Object Detection

2023-10-24 · Yan Lu, Xinzhu Ma, Lei Yang, Tianzhu Zhang 외

Geometry plays a significant role in monocular 3D object detection. It can be used to estimate object depth by using the perspective projection between object's physical size and 2D projection in the image plane, which c…

3D Object DetectionMonocular 3D Object Detectionobject-detectionObject Detection

MonoDGP: Monocular 3D Object Detection with Decoupled-Query and Geometry-Error Priors

2024-10-25 · CVPR 2025 1 · Fanqi Pu, Yifan Wang, Jiru Deng, Wenming Yang

Perspective projection has been extensively utilized in monocular 3D object detection methods. It introduces geometric priors from 2D bounding boxes and 3D object dimensions to reduce the uncertainty of depth estimation.…

3D Object DetectionDepth EstimationDepth PredictionMonocular 3D Object Detection+3