paper-with-me

홈 › Papers

On the Limits of Pseudo Ground Truth in Visual Camera Re-localisation

2021-09-01 · ICCV 2021 10 · Eric Brachmann, Martin Humenberger, Carsten Rother, Torsten Sattler

Benchmark datasets that measure camera pose accuracy have driven progress in visual re-localisation research. To obtain poses for thousands of images, it is common to use a reference algorithm to generate pseudo ground truth. Popular choices include Structure-from-Motion (SfM) and Simultaneous-Localisation-and-Mapping (SLAM) using additional sensors like depth cameras if available. Re-localisation benchmarks thus measure how well each method replicates the results of the reference algorithm. This begs the question whether the choice of the reference algorithm favours a certain family of re-localisation methods. This paper analyzes two widely used re-localisation datasets and shows that evaluation outcomes indeed vary with the choice of the reference algorithm. We thus question common beliefs in the re-localisation literature, namely that learning-based scene coordinate regression outperforms classical feature-based methods, and that RGB-D-based methods outperform RGB-based methods. We argue that any claims on ranking re-localisation methods should take the type of the reference algorithm, and the similarity of the methods to the reference algorithm, into account.

📄 PDF Abstract BibTeX arXiv:2109.00524

Code (1)

tsattler/visloc_pseudo_gt_limitations 공식 구현 pytorch

Similar Papers 제목 키워드 기반

PseudoMapTrainer: Learning Online Mapping without HD Maps

2025-08-26 · Christian Löwens, Thorben Funke, Jingchao Xie, Alexandru Paul Condurache arxiv

Online mapping models show remarkable results in predicting vectorized maps from multi-view camera images only. However, all existing approaches still rely on ground-truth high-definition maps during training, which are …

Radiologist-in-the-Loop Self-Training for Generalizable CT Metal Artifact Reduction

2025-01-26 · Chenglong Ma, Zilong Li, Yuanlin Li, Jing Han 외

Metal artifacts in computed tomography (CT) images can significantly degrade image quality and impede accurate diagnosis. Supervised metal artifact reduction (MAR) methods, trained using simulated datasets, often struggl…

Computed Tomography (CT)Metal Artifact Reduction

Long-Term Visual Localization in Dynamic Benthic Environments: A Dataset, Footprint-Based Ground Truth, and Visual Place Recognition Benchmark

2026-03-04 · Martin Kvisvik Larsen, Oscar Pizarro arxiv

Long-term visual localization has the potential to reduce cost and improve mapping quality in optical benthic monitoring with autonomous underwater vehicles (AUVs). Despite this potential, long-term visual localization i…

Visual Place RecognitionVisual Localization

Self-supervised Geometric Perception

2021-03-04 · CVPR 2021 1 · Heng Yang, Wei Dong, Luca Carlone, Vladlen Koltun

We present self-supervised geometric perception (SGP), the first general framework to learn a feature descriptor for correspondence matching without any ground-truth geometric model labels (e.g., camera poses, rigid tran…

Camera Pose EstimationPoint Cloud RegistrationPose Estimation

Using Cross-Model EgoSupervision to Learn Cooperative Basketball Intention

2017-09-05 · Gedas Bertasius, Jianbo Shi

We present a first-person method for cooperative basketball intention prediction: we predict with whom the camera wearer will cooperate in the near future from unlabeled first-person images. This is a challenging task th…

Pose Estimation