paper-with-me

홈 › Papers

Multi-View Optimization of Local Feature Geometry

2020-03-18 · ECCV 2020 8 · Mihai Dusmanu, Johannes L. Schönberger, Marc Pollefeys

In this work, we address the problem of refining the geometry of local image features from multiple views without known scene or camera geometry. Current approaches to local feature detection are inherently limited in their keypoint localization accuracy because they only operate on a single view. This limitation has a negative impact on downstream tasks such as Structure-from-Motion, where inaccurate keypoints lead to large errors in triangulation and camera localization. Our proposed method naturally complements the traditional feature extraction and matching paradigm. We first estimate local geometric transformations between tentative matches and then optimize the keypoint locations over multiple views jointly according to a non-linear least squares formulation. Throughout a variety of experiments, we show that our method consistently improves the triangulation and camera localization performance for both hand-crafted and learned local features.

📄 PDF Abstract BibTeX arXiv:2003.08348

Code (1)

mihaidusmanu/local-feature-refinement 공식 구현 pytorch

Tasks

Camera Localization

Similar Papers 제목 키워드 기반

IVGT: Implicit Visual Geometry Transformer for Neural Scene Representation

2026-05-15 · Yuqi Wu, Tianyu Hu, Wenzhao Zheng, Yuanhui Huang 외 arxiv

Reconstructing coherent 3D geometry and appearance from unposed multi-view images is a fundamental yet challenging problem in computer vision. Most existing visual geometry foundation models predict explicit geometry by …

Camera Pose EstimationNovel View Synthesis

TrianguLang: Geometry-Aware Semantic Consensus for Pose-Free 3D Localization

2026-03-09 · Bryce Grant, Aryeh Rothenberg, Atri Banerjee, Peng Wang arxiv

Localizing objects and parts from natural language in 3D space is essential for robotics, AR, and embodied AI, yet existing methods face a trade-off between the accuracy and geometric consistency of per-scene optimizatio…

MoRe: Monocular Geometry Refinement via Graph Optimization for Cross-View Consistency

2025-10-08 · Dongki Jung, Jaehoon Choi, Yonghan Lee, Sungmin Eum 외 arxiv

Monocular 3D foundation models offer an extensible solution for perception tasks, making them attractive for broader 3D vision applications. In this paper, we propose MoRe, a training-free Monocular Geometry Refinement m…

Novel View Synthesis3D Reconstruction

GPA-VGGT:Adapting VGGT to Large Scale Localization by Self-Supervised Learning with Geometry and Physics Aware Loss

2026-01-23 · Yangfan Xu, Lilian Zhang, Xiaofeng He, Pengdong Wu 외 arxiv

Transformer-based general visual geometry frameworks have shown promising performance in camera pose estimation and 3D scene understanding. Recent advancements in Visual Geometry Grounded Transformer (VGGT) models have s…

Self-Supervised LearningCamera Pose EstimationScene Understanding3D Reconstruction

MASH: Masked Anchored SpHerical Distances for 3D Shape Representation and Generation

2025-04-12 · Changhao Li, Yu Xin, Xiaowei Zhou, Ariel Shamir 외

We introduce Masked Anchored SpHerical Distances (MASH), a novel multi-view and parametrized representation of 3D shapes. Inspired by multi-view geometry and motivated by the importance of perceptual shape understanding …

3D Shape RepresentationSurface Reconstruction