paper-with-me

홈 › Papers

Split Matching for Inductive Zero-shot Semantic Segmentation

2025-05-08 · Jialei Chen, Xu Zheng, Dongyue Li, Chong Yi, Seigo Ito, Danda Pani Paudel, Luc van Gool, Hiroshi Murase, Daisuke Deguchi

Zero-shot Semantic Segmentation (ZSS) aims to segment categories that are not annotated during training. While fine-tuning vision-language models has achieved promising results, these models often overfit to seen categories due to the lack of supervision for unseen classes. As an alternative to fully supervised approaches, query-based segmentation has shown great latent in ZSS, as it enables object localization without relying on explicit labels. However, conventional Hungarian matching, a core component in query-based frameworks, needs full supervision and often misclassifies unseen categories as background in the setting of ZSS. To address this issue, we propose Split Matching (SM), a novel assignment strategy that decouples Hungarian matching into two components: one for seen classes in annotated regions and another for latent classes in unannotated regions (referred to as unseen candidates). Specifically, we partition the queries into seen and candidate groups, enabling each to be optimized independently according to its available supervision. To discover unseen candidates, we cluster CLIP dense features to generate pseudo masks and extract region-level embeddings using CLS tokens. Matching is then conducted separately for the two groups based on both class-level similarity and mask-level consistency. Additionally, we introduce a Multi-scale Feature Enhancement (MFE) module that refines decoder features through residual multi-scale aggregation, improving the model's ability to capture spatial details across resolutions. SM is the first to introduce decoupled Hungarian matching under the inductive ZSS setting, and achieves state-of-the-art performance on two standard benchmarks.

📄 PDF Abstract BibTeX arXiv:2505.05023

Code (0)

등록된 구현이 없습니다.

Tasks

Object LocalizationSemantic SegmentationZero-Shot Semantic Segmentation

Methods 이 논문이 사용한 방법론

CLIP Contrastive Language-Image Pre-training (CLIP), consisting of a simplified version of ConVIRT trained from scratch, is an efficient method of image representation learning…

Similar Papers 제목 키워드 기반

Prototypical Matching and Open Set Rejection for Zero-Shot Semantic Segmentation

2021-01-01 · ICCV 2021 10 · HUI ZHANG, Henghui Ding

The deep learning methods in addressing semantic segmentation typically demand vast amount of pixel-wise annotated training samples. In this work, we present zero-shot semantic segmentation, which aims to identify no…

SegmentationSemantic SegmentationZero-Shot Semantic Segmentation

Semantic Borrowing for Generalized Zero-Shot Learning

2021-01-30 · Xiaowei Chen

Generalized zero-shot learning (GZSL) is one of the most realistic but challenging problems due to the partiality of the classifier to supervised classes, especially under the class-inductive instance-inductive (CIII) tr…

Generalized Zero-Shot LearningMetric Learningzero-shot-classificationZero-Shot Learning

Zero-Shot Sketch-Image Hashing

2018-03-06 · CVPR 2018 6 · Yuming Shen, Li Liu, Fumin Shen, Ling Shao

Recent studies show that large-scale sketch-based image retrieval (SBIR) can be efficiently tackled by cross-modal binary representation learning methods, where Hamming distance matching significantly speeds up the proce…

Image RetrievalRepresentation LearningRetrievalSketch-Based Image Retrieval

Generative Zero-Shot Learning for Semantic Segmentation of 3D Point Clouds

2021-08-13 · Björn Michele, Alexandre Boulch, Gilles Puy, Maxime Bucher 외

While there has been a number of studies on Zero-Shot Learning (ZSL) for 2D images, its application to 3D data is still recent and scarce, with just a few methods limited to classification. We present the first generativ…

ClassificationGeneralized Zero-Shot LearningSegmentationSemantic Segmentation+1

Joint Concept Matching based Learning for Zero-Shot Recognition

2019-06-13 · Wen Tang, Ashkan Panahi, Hamid Krim

Zero-shot learning (ZSL) which aims to recognize unseen object classes by only training on seen object classes, has increasingly been of great interest in Machine Learning, and has registered with some successes. Most ex…

ObjectZero-Shot Learning