paper-with-me

홈 › Papers

FreeSeg: Unified, Universal and Open-Vocabulary Image Segmentation

2023-03-30 · CVPR 2023 1 · Jie Qin, Jie Wu, Pengxiang Yan, Ming Li, Ren Yuxi, Xuefeng Xiao, Yitong Wang, Rui Wang, Shilei Wen, Xin Pan, Xingang Wang

Recently, open-vocabulary learning has emerged to accomplish segmentation for arbitrary categories of text-based descriptions, which popularizes the segmentation system to more general-purpose application scenarios. However, existing methods devote to designing specialized architectures or parameters for specific segmentation tasks. These customized design paradigms lead to fragmentation between various segmentation tasks, thus hindering the uniformity of segmentation models. Hence in this paper, we propose FreeSeg, a generic framework to accomplish Unified, Universal and Open-Vocabulary Image Segmentation. FreeSeg optimizes an all-in-one network via one-shot training and employs the same architecture and parameters to handle diverse segmentation tasks seamlessly in the inference procedure. Additionally, adaptive prompt learning facilitates the unified model to capture task-aware and category-sensitive concepts, improving model robustness in multi-task and varied scenarios. Extensive experimental results demonstrate that FreeSeg establishes new state-of-the-art results in performance and generalization on three segmentation tasks, which outperforms the best task-specific architectures by a large margin: 5.5% mIoU on semantic segmentation, 17.6% mAP on instance segmentation, 20.1% PQ on panoptic segmentation for the unseen class on COCO.

📄 PDF Abstract BibTeX arXiv:2303.17225

Code (0)

등록된 구현이 없습니다.

Tasks

Image SegmentationInstance SegmentationOpen Vocabulary Panoptic SegmentationPanoptic SegmentationPrompt LearningSegmentationSemantic Segmentation

Similar Papers 제목 키워드 기반

FreeSeg-Diff: Training-Free Open-Vocabulary Segmentation with Diffusion Models

2024-03-29 · Barbara Toniella Corradini, Mustafa Shukor, Paul Couairon, Guillaume Couairon 외

Foundation models have exhibited unprecedented capabilities in tackling many domains and tasks. Models such as CLIP are currently widely used to bridge cross-modal representations, and text-to-image diffusion models are …

Image GenerationImage SegmentationSegmentationSemantic Segmentation+1

FreeSeg: Free Mask from Interpretable Contrastive Language-Image Pretraining for Semantic Segmentation

2022-09-27 · Yi Li, Huifeng Yao, Hualiang Wang, Xiaomeng Li

Fully supervised semantic segmentation learns from dense masks, which requires heavy annotation cost for closed set. In this paper, we use natural language as supervision without any pixel-level annotation for open world…

RetrievalSegmentationSemantic Segmentationtext similarity+1

MasQCLIP for Open-Vocabulary Universal Image Segmentation

2023-01-01 · ICCV 2023 1 · Xin Xu, Tianyi Xiong, Zheng Ding, Zhuowen Tu

We present a new method for open-vocabulary universal image segmentation, which is capable of performing instance, semantic, and panoptic segmentation under a unified framework. Our approach, called MasQCLIP, seamles…

Image SegmentationPanoptic SegmentationSegmentationSemantic Segmentation

Hierarchical Open-vocabulary Universal Image Segmentation

2023-07-03 · NeurIPS 2023 11 · Xudong Wang, Shufan Li, Konstantinos Kallidromitis, Yusuke Kato 외

Open-vocabulary image segmentation aims to partition an image into semantic regions according to arbitrary text descriptions. However, complex visual scenes can be naturally decomposed into simpler parts and abstracted a…

Image ComprehensionImage Segmentationobject-detectionObject Detection+7

OV-Uni3DETR: Towards Unified Open-Vocabulary 3D Object Detection via Cycle-Modality Propagation

2024-03-28 · Zhenyu Wang, YaLi Li, Taichi Liu, Hengshuang Zhao 외

In the current state of 3D object detection research, the severe scarcity of annotated 3D data, substantial disparities across different data modalities, and the absence of a unified architecture, have impeded the progre…

3D Object DetectionNovel Class Discoveryobject-detectionObject Detection