paper-with-me

Papers

TP3M: Transformer-based Pseudo 3D Image Matching with Reference Image

2024-05-14 · Liming Han, Zhaoxiang Liu, Shiguo Lian

Image matching is still challenging in such scenes with large viewpoints or illumination changes or with low textures. In this paper, we propose a Transformer-based pseudo 3D image matching method. It upgrades the 2D features extracted from the source image to 3D features with the help of a reference image and matches to the 2D features extracted from the destination image by the coarse-to-fine 3D matching. Our key discovery is that by introducing the reference image, the source image's fine points are screened and furtherly their feature descriptors are enriched from 2D to 3D, which improves the match performance with the destination image. Experimental results on multiple datasets show that the proposed method achieves the state-of-the-art on the tasks of homography estimation, pose estimation and visual localization especially in challenging scenes.

📄 PDF Abstract BibTeX arXiv:2405.08434

Code (0)

등록된 구현이 없습니다.

Tasks

Homography EstimationImage to 3DPose EstimationVisual Localization

Similar Papers 제목 키워드 기반

2nd Place Solution to Facebook AI Image Similarity Challenge Matching Track

2021-11-15 · SeungKee Jeon

This paper presents the 2nd place solution to the Facebook AI Image Similarity Challenge : Matching Track on DrivenData. The solution is based on self-supervised learning, and Vision Transformer(ViT). The main breaktroug…

Self-Supervised Learning

BluRef: Unsupervised Image Deblurring with Dense-Matching References

2026-03-15 · Bang-Dang Pham, Anh Tran, Cuong Pham, Minh Hoai arxiv

This paper introduces a novel unsupervised approach for image deblurring that utilizes a simple process for training data collection, thereby enhancing the applicability and effectiveness of deblurring methods. Our techn…

Image Deblurring

XoFTR: Cross-modal Feature Matching Transformer

2024-04-15 · Önder Tuzcuoğlu, Aybora Köksal, Buğra Sofu, Sinan Kalkan 외

We introduce, XoFTR, a cross-modal cross-view method for local feature matching between thermal infrared (TIR) and visible images. Unlike visible images, TIR images are less susceptible to adverse lighting and weather co…

Image Augmentation

MaDis-Stereo: Enhanced Stereo Matching via Distilled Masked Image Modeling

2024-09-04 · Jihye Ahn, Hyesong Choi, SooMin Kim, Dongbo Min

In stereo matching, CNNs have traditionally served as the predominant architectures. Although Transformer-based stereo models have been studied recently, their performance still lags behind CNN-based stereo models due to…

Depth EstimationDepth PredictionImage ReconstructionInductive Bias+1

EDGE-Shield: Efficient Denoising-staGE Shield for Violative Content Filtering via Scalable Reference-Based Matching

2026-04-04 · Takara Taniguchi, Ryohei Shimizu, Duc Minh Vo, Kota Izumi 외 arxiv

The advent of Text-to-Image generative models poses significant risks of copyright violation and deepfake generation. Since the rapid proliferation of new copyrighted works and private individuals constantly emerges, ref…

Image Generation