paper-with-me

홈 › Papers

EpiDiffVO: Geometry-Aware Epipolar Diffusion for Robust Visual Odometry

2026-05-19 · Prateeth Rao arxiv

Estimating relative pose from image pairs fundamentally requires only a minimal subset of geometrically consistent correspondences. However, most learning-based approaches rely on dense matching or direct regression, leading to redundancy and reduced geometric interpretability. In this work, we propose a sparse epipolar matching framework that predicts a compact set of correspondences optimized for geometric consistency across varying temporal baselines. To address residual noise and misalignment, we introduce an epipolar diffusion process that models correspondence uncertainty and refines keypoints toward epipolar consistency. The refined correspondences, along with depth cues, are lifted into a graph representation forming a Steiner graph that encodes relational structure between points. A graph neural network learns a compact subset of informative correspondences, which are passed to a differentiable singular value decomposition solver for end-to-end geometric estimation. Relative pose is recovered from the resulting essential matrix and evaluated in a visual odometry setting on the TartanAir and KITTI SLAM datasets. Experimental results demonstrate that combining sparse matching, diffusion-based refinement, and graph-based subset selection reduces correspondence redundancy while maintaining robust pose estimation across challenging baselines.

📄 PDF Abstract BibTeX arXiv:2605.19556

Code (0)

등록된 구현이 없습니다.

Tasks

Graph Neural NetworkPose EstimationVisual Odometry

Similar Papers 제목 키워드 기반

Epipolar Geometry Improves Video Generation Models

2025-10-24 · Orest Kupyn, Théo Uscidda, Marta Tintore Gazulla, Fabian Manhardt 외 arxiv

Video generation models have advanced significantly through the latent diffusion transformers trained with rectified flow techniques. Yet these models still struggle with geometric inconsistencies, unstable motion, and v…

Video Generation

DGSfM: Depth-Guided Scale-Aware Global Structure-from-Motion

2026-07-10 · Sithu Aung, Viktor Kocur, Yaqing Ding, Torsten Sattler 외 arxiv

Global Structure-from-Motion (SfM) is an efficient paradigm for recovering camera poses and sparse 3D structure from unordered images. However, its reliance on scale-ambiguous epipolar geometry makes global positioning s…

ProDiG: Progressive Diffusion-Guided Gaussian Splatting for Aerial to Ground Reconstruction

2026-04-02 · Sirshapan Mitra, Yogesh S. Rawat arxiv

Generating ground-level views and coherent 3D site models from aerial-only imagery is challenging due to extreme viewpoint changes, missing intermediate observations, and large scale variations. Existing methods either r…

CamPVG: Camera-Controlled Panoramic Video Generation with Epipolar-Aware Diffusion

2025-09-24 · Chenhao Ji, Chaohui Yu, Junyao Gao, Fan Wang 외 arxiv

Recently, camera-controlled video generation has seen rapid development, offering more precise control over video generation. However, existing methods predominantly focus on camera control in perspective projection vide…

Video Generation

Rolling Shutter Camera Relative Pose: Generalized Epipolar Geometry

2016-05-02 · CVPR 2016 6 · Yuchao Dai, Hongdong Li, Laurent Kneip

The vast majority of modern consumer-grade cameras employ a rolling shutter mechanism. In dynamic geometric computer vision applications such as visual SLAM, the so-called rolling shutter effect therefore needs to be pro…