Multilevel Saliency-Guided Self-Supervised Learning for Image Anomaly Detection
Anomaly detection (AD) is a fundamental task in computer vision. It aims to identify incorrect image data patterns which deviate from the normal ones. Conventional methods generally address AD by preparing augmented negative samples to enforce self-supervised learning. However, these techniques typically do not consider semantics during augmentation, leading to the generation of unrealistic or invalid negative samples. Consequently, the feature extraction network can be hindered from embedding critical features. In this study, inspired by visual attention learning approaches, we propose CutSwap, which leverages saliency guidance to incorporate semantic cues for augmentation. Specifically, we first employ LayerCAM to extract multilevel image features as saliency maps and then perform clustering to obtain multiple centroids. To fully exploit saliency guidance, on each map, we select a pixel pair from the cluster with the highest centroid saliency to form a patch pair. Such a patch pair includes highly similar context information with dense semantic correlations. The resulting negative sample is created by swapping the locations of the patch pair. Compared to prior augmentation methods, CutSwap generates more subtle yet realistic negative samples to facilitate quality feature learning. Extensive experimental and ablative evaluations demonstrate that our method achieves state-of-the-art AD performance on two mainstream AD benchmark datasets.
Code (0)
등록된 구현이 없습니다.
Tasks
Anomaly DetectionSelf-Supervised LearningSimilar Papers 제목 키워드 기반
SSiT: Saliency-guided Self-supervised Image Transformer for Diabetic Retinopathy Grading
Self-supervised Learning (SSL) has been widely applied to learn image representations through exploiting unlabeled images. However, it has not been fully explored in the medical image analysis field. In this work, Salien…
Contrastive LearningDiabetic Retinopathy GradingMedical Image AnalysisSelf-Supervised LearningSaliency guided deep network for weakly-supervised image segmentation
Weakly-supervised image segmentation is an important task in computer vision. A key problem is how to obtain high quality objects location from image-level category. Classification activation mapping is a common method w…
Image SegmentationSegmentationSemantic SegmentationSaliency Guided Contrastive Learning on Scene Images
Self-supervised learning holds promise in leveraging large numbers of unlabeled data. However, its success heavily relies on the highly-curated dataset, e.g., ImageNet, which still needs human cleaning. Directly learning…
Contrastive LearningLinear evaluationRepresentation LearningSelf-Supervised LearningSaliency Guided Self-attention Network for Weakly and Semi-supervised Semantic Segmentation
Weakly supervised semantic segmentation (WSSS) using only image-level labels can greatly reduce the annotation cost and therefore has attracted considerable research interest. However, its performance is still inferior t…
SegmentationSemantic SegmentationSemi-Supervised Semantic SegmentationWeakly supervised Semantic Segmentation+1Attention-Guided Lidar Segmentation and Odometry Using Image-to-Point Cloud Saliency Transfer
LiDAR odometry estimation and 3D semantic segmentation are crucial for autonomous driving, which has achieved remarkable advances recently. However, these tasks are challenging due to the imbalance of points in different…
3D Semantic SegmentationAutonomous DrivingSegmentationSemantic Segmentation+1