paper-with-me

홈 › Papers

Robust Scene Coordinate Regression via Geometrically-Consistent Global Descriptors

2025-12-19 · Son Tung Nguyen, Alejandro Fontan, Michael Milford, Tobias Fischer arxiv

Recent learning-based visual localization methods use global descriptors to disambiguate visually similar places, but existing approaches often derive these descriptors from geometric cues alone (e.g., covisibility graphs), limiting their discriminative power and reducing robustness in the presence of noisy geometric constraints. We propose an aggregator module that learns global descriptors consistent with both geometrical structure and visual similarity, ensuring that images are close in descriptor space only when they are visually similar and spatially connected. This corrects erroneous associations caused by unreliable overlap scores. Using a batch-mining strategy based solely on the overlap scores and a modified contrastive loss, our method trains without manual place labels and generalizes across diverse environments. Experiments on challenging benchmarks show substantial localization gains in large-scale environments while preserving computational and memory efficiency. Code is available at https://github.com/sontung/robust_scr.

📄 PDF Abstract BibTeX arXiv:2512.17226

Code (0)

등록된 구현이 없습니다.

Tasks

Visual Localization

Similar Papers 제목 키워드 기반

POMA-3D: The Point Map Way to 3D Scene Understanding

2025-11-20 · Ye Mao, Weixun Luo, Ranran Huang, Junpeng Jing 외 arxiv

In this paper, we introduce POMA-3D, the first self-supervised 3D representation model learned from point maps. Point maps encode explicit 3D coordinates on a structured 2D grid, preserving global 3D geometry while remai…

Representation LearningScene UnderstandingQuestion Answering

Large Scale Joint Semantic Re-Localisation and Scene Understanding via Globally Unique Instance Coordinate Regression

2019-09-23 · Ignas Budvytis, Marvin Teichmann, Tomas Vojir, Roberto Cipolla

In this work we present a novel approach to joint semantic localisation and scene understanding. Our work is motivated by the need for localisation algorithms which not only predict 6-DoF camera pose but also simultaneou…

3D geometryAutonomous DrivingCamera Pose EstimationPose Estimation+2

Full-Frame Scene Coordinate Regression for Image-Based Localization

2018-02-09 · Xiaotian Li, Juha Ylioinas, Juho Kannala

Image-based localization, or camera relocalization, is a fundamental problem in computer vision and robotics, and it refers to estimating camera pose from an image. Recent state-of-the-art approaches use learning based m…

Camera RelocalizationData AugmentationDecoderImage-Based Localization+1

SyncNoise: Geometrically Consistent Noise Prediction for Text-based 3D Scene Editing

2024-06-25 · Ruihuang Li, Liyi Chen, Zhengqiang Zhang, Varun Jampani 외

Text-based 2D diffusion models have demonstrated impressive capabilities in image generation and editing. Meanwhile, the 2D diffusion models also exhibit substantial potentials for 3D editing tasks. However, how to achie…

3D scene EditingImage Generation

KitchenTwin: Semantically and Geometrically Grounded 3D Kitchen Digital Twins

2026-03-25 · Quanyun Wu, Kyle Gao, Daniel Long, David A. Clausi 외 arxiv

Embodied AI training and evaluation require object-centric digital twin environments with accurate metric geometry and semantic grounding. Recent transformer-based feedforward reconstruction methods can efficiently predi…

Point Clouds