paper-with-me

Papers

Self-supervised Semantic Segmentation Grounded in Visual Concepts

2022-03-25 · Wenbin He, William Surmeier, Arvind Kumar Shekar, Liang Gou, Liu Ren

Unsupervised semantic segmentation requires assigning a label to every pixel without any human annotations. Despite recent advances in self-supervised representation learning for individual images, unsupervised semantic segmentation with pixel-level representations is still a challenging task and remains underexplored. In this work, we propose a self-supervised pixel representation learning method for semantic segmentation by using visual concepts (i.e., groups of pixels with semantic meanings, such as parts, objects, and scenes) extracted from images. To guide self-supervised learning, we leverage three types of relationships between pixels and concepts, including the relationships between pixels and local concepts, local and global concepts, as well as the co-occurrence of concepts. We evaluate the learned pixel embeddings and visual concepts on three datasets, including PASCAL VOC 2012, COCO 2017, and DAVIS 2017. Our results show that the proposed method gains consistent and substantial improvements over recent unsupervised semantic segmentation approaches, and also demonstrate that visual concepts can reveal insights into image datasets.

📄 PDF Abstract BibTeX arXiv:2203.13868

Code (0)

등록된 구현이 없습니다.

Tasks

Representation LearningSegmentationSelf-Supervised LearningSemantic SegmentationUnsupervised Semantic Segmentation

Similar Papers 제목 키워드 기반

DatUS^2: Data-driven Unsupervised Semantic Segmentation with Pre-trained Self-supervised Vision Transformer

2024-01-23 · Sonal Kumar, Arijit Sur, Rashmi Dutta Baruah

Successive proposals of several self-supervised training schemes continue to emerge, taking one step closer to developing a universal foundation model. In this process, the unsupervised downstream tasks are recognized as…

SegmentationSemantic SegmentationUnsupervised Semantic Segmentation

Understanding Self-Supervised Features for Learning Unsupervised Instance Segmentation

2023-11-24 · Paul Engstler, Luke Melas-Kyriazi, Christian Rupprecht, Iro Laina

Self-supervised learning (SSL) can be used to solve complex visual tasks without human labels. Self-supervised representations encode useful semantic information about images, and as a result, they have already been used…

Instance SegmentationSegmentationSelf-Supervised LearningSemantic Segmentation+3

Word Discovery in Visually Grounded, Self-Supervised Speech Models

2022-03-28 · Puyuan Peng, David Harwath

We present a method for visually-grounded spoken term discovery. After training either a HuBERT or wav2vec2.0 model to associate spoken captions with natural images, we show that powerful word segmentation and clustering…

ClusteringSegmentationVisual Grounding

Self-Supervised Difference Detection for Weakly-Supervised Semantic Segmentation

2019-11-04 · ICCV 2019 10 · Wataru Shimoda, Keiji Yanai

To minimize the annotation costs associated with the training of semantic segmentation models, researchers have extensively investigated weakly-supervised segmentation approaches. In the current weakly-supervised segment…

SegmentationSemantic SegmentationWeakly supervised segmentationWeakly supervised Semantic Segmentation+1

Syllable Discovery and Cross-Lingual Generalization in a Visually Grounded, Self-Supervised Speech Model

2023-05-19 · Puyuan Peng, Shang-Wen Li, Okko Räsänen, Abdelrahman Mohamed 외

In this paper, we show that representations capturing syllabic units emerge when training a self-supervised speech model with a visually-grounded training objective. We demonstrate that a nearly identical model architect…

Language ModelingLanguage ModellingMasked Language ModelingSegmentation+2