paper-with-me

홈 › Papers

Semantic Graph Consistency: Going Beyond Patches for Regularizing Self-Supervised Vision Transformers

2024-06-18 · Chaitanya Devaguptapu, Sumukh Aithal, Shrinivas Ramasubramanian, Moyuru Yamada, Manohar Kaul

Self-supervised learning (SSL) with vision transformers (ViTs) has proven effective for representation learning as demonstrated by the impressive performance on various downstream tasks. Despite these successes, existing ViT-based SSL architectures do not fully exploit the ViT backbone, particularly the patch tokens of the ViT. In this paper, we introduce a novel Semantic Graph Consistency (SGC) module to regularize ViT-based SSL methods and leverage patch tokens effectively. We reconceptualize images as graphs, with image patches as nodes and infuse relational inductive biases by explicit message passing using Graph Neural Networks into the SSL framework. Our SGC loss acts as a regularizer, leveraging the underexploited patch tokens of ViTs to construct a graph and enforcing consistency between graph features across multiple views of an image. Extensive experiments on various datasets including ImageNet, RESISC and Food-101 show that our approach significantly improves the quality of learned representations, resulting in a 5-10\% increase in performance when limited labeled data is used for linear evaluation. These experiments coupled with a comprehensive set of ablations demonstrate the promise of our approach in various settings.

📄 PDF Abstract BibTeX arXiv:2406.12944

Code (0)

등록된 구현이 없습니다.

Tasks

Linear evaluationRepresentation LearningSelf-Supervised Learning

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

Structure-aware Semantic Discrepancy and Consistency for 3D Medical Image Self-supervised Learning

2025-07-03 · Tan Pan, Zhaorui Tan, Kaiyu Guo, Dongli Xu 외 arxiv

3D medical image self-supervised learning (mSSL) holds great promise for medical analysis. Effectively supporting broader applications requires considering anatomical structure variations in location, scale, and morpholo…

Self-Supervised Learning

Importance of Self-Consistency in Active Learning for Semantic Segmentation

2020-08-04 · S. Alireza Golestaneh, Kris M. Kitani

We address the task of active learning in the context of semantic segmentation and show that self-consistency can be a powerful source of self-supervision to greatly improve the performance of a data-driven model with ac…

Active LearningSegmentationSemantic Segmentation

Leveraging Hidden Positives for Unsupervised Semantic Segmentation

2023-03-27 · CVPR 2023 1 · Hyun Seok Seong, WonJun Moon, SuBeen Lee, Jae-Pil Heo

Dramatic demand for manpower to label pixel-level annotations triggered the advent of unsupervised semantic segmentation. Although the recent work employing the vision transformer (ViT) backbone shows exceptional perform…

Contrastive LearningSemantic SegmentationUnsupervised Semantic Segmentation

Beyond Patches: Mining Interpretable Part-Prototypes for Explainable AI

2025-04-16 · Mahdi Alehdaghi, Rajarshi Bhattacharya, Pourya Shamsolmoali, Rafael M. O. Cruz 외

Deep learning has provided considerable advancements for multimedia systems, yet the interpretability of deep models remains a challenge. State-of-the-art post-hoc explainability methods, such as GradCAM, provide visual …

Unsupervised Part Discovery

SAM-CP: Marrying SAM with Composable Prompts for Versatile Segmentation

2024-07-23 · Pengfei Chen, Lingxi Xie, Xinyue Huo, Xuehui Yu 외

The Segment Anything model (SAM) has shown a generalized ability to group image pixels into patches, but applying it to semantic-aware segmentation still faces major challenges. This paper presents SAM-CP, a simple appro…

Panoptic SegmentationSegmentation