paper-with-me

홈 › Papers

Matcher: Segment Anything with One Shot Using All-Purpose Feature Matching

2023-05-22 · Yang Liu, Muzhi Zhu, Hengtao Li, Hao Chen, Xinlong Wang, Chunhua Shen

Powered by large-scale pre-training, vision foundation models exhibit significant potential in open-world image understanding. However, unlike large language models that excel at directly tackling various language tasks, vision foundation models require a task-specific model structure followed by fine-tuning on specific tasks. In this work, we present Matcher, a novel perception paradigm that utilizes off-the-shelf vision foundation models to address various perception tasks. Matcher can segment anything by using an in-context example without training. Additionally, we design three effective components within the Matcher framework to collaborate with these foundation models and unleash their full potential in diverse perception tasks. Matcher demonstrates impressive generalization performance across various segmentation tasks, all without training. For example, it achieves 52.7% mIoU on COCO-20$^i$ with one example, surpassing the state-of-the-art specialist model by 1.6%. In addition, Matcher achieves 33.0% mIoU on the proposed LVIS-92$^i$ for one-shot semantic segmentation, outperforming the state-of-the-art generalist model by 14.4%. Our visualization results further showcase the open-world generality and flexibility of Matcher when applied to images in the wild. Our code can be found at https://github.com/aim-uofa/Matcher.

📄 PDF Abstract BibTeX arXiv:2305.13310

Code (1)

aim-uofa/matcher 공식 구현 pytorch

Tasks

AllFew-Shot Semantic SegmentationSegmentationSemantic Segmentation

Similar Papers 제목 키워드 기반

SAMatcher: Co-Visibility Modeling with Segment Anything for Robust Feature Matching

2026-06-02 · Xu Pan, Qiyuan Ma, Mingyue Dong, He Chen 외 arxiv

Reliable correspondence estimation is a fundamental problem in image processing, underpinning applications such as Structure from Motion, visual localization, and image registration. Existing learning-based methods have …

Representation LearningVisual LocalizationImage RegistrationImage Matching

SANSA: Unleashing the Hidden Semantics in SAM2 for Few-Shot Segmentation

2025-05-27 · Claudia Cuttano, Gabriele Trivigno, Giuseppe Averta, Carlo Masone

Few-shot segmentation aims to segment unseen object categories from just a handful of annotated examples. This requires mechanisms that can both identify semantically related objects across images and accurately produce …

Object TrackingSegmentation

Are Pretrained Image Matchers Good Enough for SAR-Optical Satellite Registration?

2026-04-11 · Isaac Corley, Alex Stoken, Gabriele Berton arxiv

Cross-modal optical-SAR (Synthetic Aperture Radar) registration is a bottleneck for disaster-response via remote sensing, yet modern image matchers are developed and benchmarked almost exclusively on natural-image domain…

Domain AdaptationImage Matching

Zero-Shot Polygon Matching with Pre-trained Models for Pose Estimation and Polygon Cloud from Challenging Stereo

2025-11-08 · Chang Li, Xingtao Peng arxiv

While stereo matching has achieved maturity for 0D point and 1D line primitives, establishing correspondences for 2D polygons remains largely unexplored due to challenges including disparity discontinuity, scale variatio…

Zero-shot Generalization3D ReconstructionPose Estimation

SAM.MD: Zero-shot medical image segmentation capabilities of the Segment Anything Model

2023-04-10 · Saikat Roy, Tassilo Wald, Gregor Koehler, Maximilian R. Rokuss 외

Foundation models have taken over natural language processing and image generation domains due to the flexibility of prompting. With the recent introduction of the Segment Anything Model (SAM), this prompt-driven paradig…

Image GenerationImage SegmentationMedical Image SegmentationOrgan Segmentation+3