paper-with-me

홈 › Papers

Is This The Right Place? Geometric-Semantic Pose Verification for Indoor Visual Localization

2019-08-13 · ICCV 2019 10 · Hajime Taira, Ignacio Rocco, Jiri Sedlar, Masatoshi Okutomi, Josef Sivic, Tomas Pajdla, Torsten Sattler, Akihiko Torii

Visual localization in large and complex indoor scenes, dominated by weakly textured rooms and repeating geometric patterns, is a challenging problem with high practical relevance for applications such as Augmented Reality and robotics. To handle the ambiguities arising in this scenario, a common strategy is, first, to generate multiple estimates for the camera pose from which a given query image was taken. The pose with the largest geometric consistency with the query image, e.g., in the form of an inlier count, is then selected in a second stage. While a significant amount of research has concentrated on the first stage, there is considerably less work on the second stage. In this paper, we thus focus on pose verification. We show that combining different modalities, namely appearance, geometry, and semantics, considerably boosts pose verification and consequently pose accuracy. We develop multiple hand-crafted as well as a trainable approach to join into the geometric-semantic verification and show significant improvements over state-of-the-art on a very challenging indoor dataset.

📄 PDF Abstract BibTeX arXiv:1908.04598

Code (0)

등록된 구현이 없습니다.

Tasks

Visual Localization

Similar Papers 제목 키워드 기반

Geometry-Aware Localized Watermarking for Copyright Protection in Embedding-as-a-Service

2026-04-13 · Zhimin Chen, Xiaojie Liang, Wenbo Xu, Yuxuan Liu 외 arxiv

Embedding-as-a-Service (EaaS) has become an important semantic infrastructure for natural language and multimedia applications, but it is highly vulnerable to model stealing and copyright infringement. Existing EaaS wate…

COMPASS: COmpact Multi-channel Prior-map And Scene Signature for Floor-Plan-Based Visual Localization

2026-04-28 · Muhammad Shaheer, Miguel Fernandez-Cortizas, Asier Bikandi-Noya, Holger Voos 외 arxiv

Architectural floor plans are widely available priors which contain not only geometry but also the semantic information of the environment, yet existing localization methods largely ignore this semantic information. To a…

Visual Localization

GV-Bench: Benchmarking Local Feature Matching for Geometric Verification of Long-term Loop Closure Detection

2024-07-16 · Jingwen Yu, Hanjing Ye, Jianhao Jiao, Ping Tan 외

Visual loop closure detection is an important module in visual simultaneous localization and mapping (SLAM), which associates current camera observation with previously visited places. Loop closures correct drifts in tra…

BenchmarkingLoop Closure DetectionPose EstimationSimultaneous Localization and Mapping+1

Visual Attention Reasoning via Hierarchical Search and Self-Verification

2025-10-21 · Wei Cai, Jian Zhao, Yuchen Yuan, Tianle Zhang 외 arxiv

Multimodal Large Language Models (MLLMs) frequently hallucinate due to their reliance on fragile, linear reasoning and weak visual grounding. We propose Visual Attention Reasoning (VAR), a reinforcement learning framewor…

Reinforcement LearningVisual Grounding

FirePlace: Geometric Refinements of LLM Common Sense Reasoning for 3D Object Placement

2025-01-01 · CVPR 2025 1 · IAn Huang, Yanan Bao, Karen Truong, Howard Zhou 외

Scene generation with 3D assets presents a complex challenge, requiring both high-level semantic understanding and low-level geometric reasoning. While Multimodal Large Language Models (MLLMs) excel at semantic tasks…

3D geometryCommon Sense ReasoningScene Generation