paper-with-me

Papers

Mismatched: Evaluating the Limits of Image Matching Approaches and Benchmarks

2024-08-29 · Sierra Bonilla, Chiara Di Vece, Rema Daher, Xinwei Ju, Danail Stoyanov, Francisco Vasconcelos, Sophia Bano

Three-dimensional (3D) reconstruction from two-dimensional images is an active research field in computer vision, with applications ranging from navigation and object tracking to segmentation and three-dimensional modeling. Traditionally, parametric techniques have been employed for this task. However, recent advancements have seen a shift towards learning-based methods. Given the rapid pace of research and the frequent introduction of new image matching methods, it is essential to evaluate them. In this paper, we present a comprehensive evaluation of various image matching methods using a structure-from-motion pipeline. We assess the performance of these methods on both in-domain and out-of-domain datasets, identifying key limitations in both the methods and benchmarks. We also investigate the impact of edge detection as a pre-processing step. Our analysis reveals that image matching for 3D reconstruction remains an open challenge, necessitating careful selection and tuning of models for specific scenarios, while also highlighting mismatches in how metrics currently represent method performance.

📄 PDF Abstract BibTeX arXiv:2408.16445

Code (1)

surgical-vision/colmap-match-converter 공식 구현 pytorch

Tasks

3D ReconstructionEdge DetectionObject Tracking

Similar Papers 제목 키워드 기반

Grounded Image Text Matching with Mismatched Relation Reasoning

2023-08-02 · ICCV 2023 1 · Yu Wu, Yana Wei, Haozhe Wang, Yongfei Liu 외

This paper introduces Grounded Image Text Matching with Mismatched Relation (GITM-MR), a novel visual-linguistic joint task that evaluates the relation understanding capabilities of transformer-based pre-trained models. …

Image-text matchingRelationSentenceText Matching

Negative-Aware Attention Framework for Image-Text Matching

2022-01-01 · CVPR 2022 1 · Kun Zhang, Zhendong Mao, Quan Wang, Yongdong Zhang

Image-text matching, as a fundamental task, bridges the gap between vision and language. The key of this task is to accurately measure similarity between these two modalities. Prior work measuring this similarity mai…

Image-text matchingText Matchingtext similarity

Self-Supervised Visual Acoustic Matching

2023-07-27 · NeurIPS 2023 11

Acoustic matching aims to re-synthesize an audio clip to sound as if it were recorded in a target acoustic environment. Existing methods assume access to paired training data, where the audio is observed in both source a…

Diversity

SIGMA: Semantic-complete Graph Matching for Domain Adaptive Object Detection

2022-03-12 · CVPR 2022 1 · Wuyang Li, Xinyu Liu, Yixuan Yuan

Domain Adaptive Object Detection (DAOD) leverages a labeled domain to learn an object detector generalizing to a novel domain free of annotations. Recent advances align class-conditional distributions by narrowing down c…

Graph MatchingHallucinationobject-detectionObject Detection

ManiGAN: Text-Guided Image Manipulation

2019-12-12 · Bowen Li, Xiaojuan Qi, Thomas Lukasiewicz, Philip H. S. Torr

The goal of our paper is to semantically edit parts of an image matching a given text that describes desired attributes (e.g., texture, colour, and background), while preserving other contents that are irrelevant to the …

Generative Adversarial NetworkImage ManipulationText-to-Image Generation