paper-with-me

Papers

Doppelgangers++: Improved Visual Disambiguation with Geometric 3D Features

2024-12-08 · CVPR 2025 1 · Yuanbo Xiangli, Ruojin Cai, HanYu Chen, Jeffrey Byrne, Noah Snavely

Accurate 3D reconstruction is frequently hindered by visual aliasing, where visually similar but distinct surfaces (aka, doppelgangers), are incorrectly matched. These spurious matches distort the structure-from-motion (SfM) process, leading to misplaced model elements and reduced accuracy. Prior efforts addressed this with CNN classifiers trained on curated datasets, but these approaches struggle to generalize across diverse real-world scenes and can require extensive parameter tuning. In this work, we present Doppelgangers++, a method to enhance doppelganger detection and improve 3D reconstruction accuracy. Our contributions include a diversified training dataset that incorporates geo-tagged images from everyday scenes to expand robustness beyond landmark-based datasets. We further propose a Transformer-based classifier that leverages 3D-aware features from the MASt3R model, achieving superior precision and recall across both in-domain and out-of-domain tests. Doppelgangers++ integrates seamlessly into standard SfM and MASt3R-SfM pipelines, offering efficiency and adaptability across varied scenes. To evaluate SfM accuracy, we introduce an automated, geotag-based method for validating reconstructed models, eliminating the need for manual inspection. Through extensive experiments, we demonstrate that Doppelgangers++ significantly enhances pairwise visual disambiguation and improves 3D reconstruction quality in complex and diverse scenarios.

📄 PDF Abstract BibTeX arXiv:2412.05826

Code (0)

등록된 구현이 없습니다.

Tasks

3D Reconstruction

Similar Papers 제목 키워드 기반

Doppelgangers: Learning to Disambiguate Images of Similar Structures

2023-09-05 · ICCV 2023 1 · Ruojin Cai, Joseph Tung, Qianqian Wang, Hadar Averbuch-Elor 외

We consider the visual disambiguation task of determining whether a pair of visually similar images depict the same or distinct 3D surfaces (e.g., the same or opposite sides of a symmetric building). Illusory image match…

3D ReconstructionBinary Classification

Doppelgangers and Adversarial Vulnerability

2025-01-01 · CVPR 2025 1 · George Kamberov

Many machine learning (ML) classifiers are claimed to outperform humans, but they still make mistakes that humans do not. The most notorious examples of such mistakes are adversarial visual metamers. This paper aims …

DisCo-FLoc: Semantic-Free Floorplan Localization via $SE(2)$-Aware Contrastive Disambiguation

2026-01-05 · Ping Zhong, Shiyong Meng, Bolei Chen, Tao Zou 외 arxiv

Visual Floorplan Localization (FLoc) struggles with severe structural aliasing caused by repetitive minimalist layouts. This occurs because physically distant poses share highly similar visual-geometric features, which d…

Contrastive Learning

ENEAS: Embedding-guided Neural Ensemble for Adaptive Segmentation

2026-09-03 · Javier del Pino, Salvador Rodríguez, Alejandro Garabito, Javier Álvarez 외 hf

We present ENEAS, a unified, text-promptable method for instance tracking and semantic discovery. Text-promptable segmentation models, including the latest foundation models such as SAM 3, still suffer from temporal hall…

3D Reconstruction

Unsupervised Visual Sense Disambiguation for Verbs using Multimodal Embeddings

2016-03-30 · NAACL 2016 6 · Spandana Gella, Mirella Lapata, Frank Keller

We introduce a new task, visual sense disambiguation for verbs: given an image and a verb, assign the correct sense of the verb, i.e., the one that describes the action depicted in the image. Just as textual word sense d…

Image DescriptionImage RetrievalRetrievalWord Sense Disambiguation