paper-with-me

홈 › Papers

Associating Spatially-Consistent Grouping with Text-supervised Semantic Segmentation

2023-04-03 · Yabo Zhang, ZiHao Wang, Jun Hao Liew, Jingjia Huang, Manyu Zhu, Jiashi Feng, WangMeng Zuo

In this work, we investigate performing semantic segmentation solely through the training on image-sentence pairs. Due to the lack of dense annotations, existing text-supervised methods can only learn to group an image into semantic regions via pixel-insensitive feedback. As a result, their grouped results are coarse and often contain small spurious regions, limiting the upper-bound performance of segmentation. On the other hand, we observe that grouped results from self-supervised models are more semantically consistent and break the bottleneck of existing methods. Motivated by this, we introduce associate self-supervised spatially-consistent grouping with text-supervised semantic segmentation. Considering the part-like grouped results, we further adapt a text-supervised model from image-level to region-level recognition with two core designs. First, we encourage fine-grained alignment with a one-way noun-to-region contrastive loss, which reduces the mismatched noun-region pairs. Second, we adopt a contextually aware masking strategy to enable simultaneous recognition of all grouped regions. Coupled with spatially-consistent grouping and region-adapted recognition, our method achieves 59.2% mIoU and 32.4% mIoU on Pascal VOC and Pascal Context benchmarks, significantly surpassing the state-of-the-art methods.

📄 PDF Abstract BibTeX arXiv:2304.01114

Code (0)

등록된 구현이 없습니다.

Tasks

SegmentationSemantic SegmentationSentence

Methods 이 논문이 사용한 방법론

AWARE We propose to theoretically and empirically examine the effect of incorporating weighting schemes into walk-aggregating GNNs. To this end, we propose a simple, interpretable, and…

Similar Papers 제목 키워드 기반

ViFiCon: Vision and Wireless Association Via Self-Supervised Contrastive Learning

2022-10-11 · Nicholas Meegan, Hansi Liu, Bryan Cao, Abrar Alali 외

We introduce ViFiCon, a self-supervised contrastive learning scheme which uses synchronized information across vision and wireless modalities to perform cross-modal association. Specifically, the system uses pedestrian d…

Contrastive LearningRegion Proposal

Unsupervised Hierarchical Semantic Segmentation with Multiview Cosegmentation and Clustering Transformers

2022-04-25 · CVPR 2022 1 · Tsung-Wei Ke, Jyh-Jing Hwang, Yunhui Guo, Xudong Wang 외

Unsupervised semantic segmentation aims to discover groupings within and across images that capture object and view-invariance of a category without external supervision. Grouping naturally has levels of granularity, cre…

ClusteringSegmentationSemantic SegmentationUnsupervised Semantic Segmentation

A Spatial Guided Self-supervised Clustering Network for Medical Image Segmentation

2021-07-11 · Euijoon Ahn, Dagan Feng, Jinman Kim

The segmentation of medical images is a fundamental step in automated clinical decision support systems. Existing medical image segmentation methods based on supervised deep learning, however, remain problematic because …

ClusteringImage SegmentationMedical Image SegmentationSegmentation+1

Space-Time Forecasting of Dynamic Scenes with Motion-aware Gaussian Grouping

2026-02-25 · Junmyeong Lee, Hoseung Choi, Minsu Cho arxiv

Forecasting dynamic scenes remains a fundamental challenge in computer vision, as limited observations make it difficult to capture coherent object-level motion and long-term temporal evolution. We present Motion Group-a…

Foundation AI Models for Aerosol Optical Depth Estimation from PACE Satellite Data

2026-05-01 · Zahid Hassan Tushar, Sanjay Purushotham arxiv

Aerosol Optical Depth (AOD) retrieval is essential for Earth observation, supporting applications from air quality monitoring to climate studies. Conventional physics-based AOD retrieval methods formulate the problem as …

Depth Estimation