Benchmarking Robust Self-Supervised Learning Across Diverse Downstream Tasks
Large-scale vision models have become integral in many applications due to their unprecedented performance and versatility across downstream tasks. However, the robustness of these foundation models has primarily been explored for a single task, namely image classification. The vulnerability of other common vision tasks, such as semantic segmentation and depth estimation, remains largely unknown. We present a comprehensive empirical evaluation of the adversarial robustness of self-supervised vision encoders across multiple downstream tasks. Our attacks operate in the encoder embedding space and at the downstream task output level. In both cases, current state-of-the-art adversarial fine-tuning techniques tested only for classification significantly degrade clean and robust performance on other tasks. Since the purpose of a foundation model is to cater to multiple applications at once, our findings reveal the need to enhance encoder robustness more broadly. Our code is available at ${github.com/layer6ai-labs/ssl-robustness}$.
Code (1)
Tasks
Adversarial RobustnessBenchmarkingDepth Estimationimage-classificationImage ClassificationSelf-Supervised LearningSemantic SegmentationSimilar Papers 제목 키워드 기반
RoFt-Mol: Benchmarking Robust Fine-Tuning with Molecular Graph Foundation Models
In the era of foundation models, fine-tuning pre-trained models for specific downstream tasks has become crucial. This drives the need for robust fine-tuning methods to address challenges such as model overfitting and sp…
Benchmarking Self-Supervised Contrastive Learning Methods for Image-Based Plant Phenotyping
The rise of self-supervised learning (SSL) methods in recent years presents an opportunity to leverage unlabeled and domain-specific datasets generated by image-based plant phenotyping platforms to accelerate plant breed…
BenchmarkingContrastive LearningHead DetectionPlant Phenotyping+1Benchmarking Self-Supervised Learning on Diverse Pathology Datasets
Computational pathology can lead to saving human lives, but models are annotation hungry and pathology images are notoriously expensive to annotate. Self-supervised learning has shown to be an effective method for utiliz…
BenchmarkingClassificationInstance SegmentationSelf-Supervised Learning+1Speech Self-Supervised Representation Benchmarking: Are We Doing it Right?
Self-supervised learning (SSL) has recently allowed leveraging large datasets of unlabeled speech signals to reach impressive performance on speech tasks using only small amounts of annotated data. The high number of pro…
BenchmarkingDecoderSelf-Supervised LearningSpeech Self-Supervised Representations Benchmarking: a Case for Larger Probing Heads
Self-supervised learning (SSL) leverages large datasets of unlabeled speech to reach impressive performance with reduced amounts of annotated data. The high number of proposed approaches fostered the emergence of compreh…
BenchmarkingSelf-Supervised Learning