Invariant and consistent: Unsupervised representation learning for few-shot visual recognition
Few-shot visual recognition aims to identify novel unseen classes with few labels while learning generalized prior knowledge from base classes. Recent ideas propose to explore this problem in an unsupervised setting, i.e., without any labels in base classes, which reduces the heavy consumption of manual annotations. In this paper, we build upon a self-supervised insight and propose a novel unsupervised learning approach that joints Invariant and Consistent (InCo) representation for the few shot task. For the invariant representation operation, we present a geometric invariance module to construct the rotation prediction of each instance, which learns the intra-instance variance and improves the feature discrimination. To further build consistency representation of inter-instance, we propose a pairwise consistency module from two contrastive learning aspects: a holistic contrastive learning with historical training queues, and a local contrastive learning for enhancing the representation of current training samples. Moreover, to better facilitate contrastive learning among features, we introduce an asymmetric convolutional architecture to encode high-quality representations. Comprehensive experiments on 4 public benchmarks demonstrate the utility of our approach and the superiority compared to existing approaches.
Code (0)
등록된 구현이 없습니다.
Tasks
Contrastive LearningRepresentation LearningUnsupervised Few-Shot Image ClassificationMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
A One-Shot Texture-Perceiving Generative Adversarial Network for Unsupervised Surface Inspection
Visual surface inspection is a challenging task owing to the highly diverse appearance of target surfaces and defective regions. Previous attempts heavily rely on vast quantities of training examples with manual annotati…
Generative Adversarial NetworkPushing the Limits of Fewshot Anomaly Detection in Industry Vision: Graphcore
In the area of fewshot anomaly detection (FSAD), efficient visual feature plays an essential role in memory bank M-based methods. However, these methods do not account for the relationship between the visual feature and …
Anomaly DetectionUnifying Unsupervised Domain Adaptation and Zero-Shot Visual Recognition
Unsupervised domain adaptation aims to transfer knowledge from a source domain to a target domain so that the target domain data can be recognized without any explicit labelling information for this domain. One limitatio…
Domain Adaptationdomain classificationGeneralized Zero-Shot LearningUnsupervised Domain Adaptation+1Unsupervised Learning of Invariant Representations in Hierarchical Architectures
The present phase of Machine Learning is characterized by supervised learning algorithms relying on large sets of labeled examples ($n \to \infty$). The next phase is likely to focus on algorithms capable of learning fro…
Object Recognitionspeech-recognitionSpeech RecognitionLearning Causal Domain-Invariant Temporal Dynamics for Few-Shot Action Recognition
Few-shot action recognition aims at quickly adapting a pre-trained model to the novel data with a distribution shift using only a limited number of samples. Key challenges include how to identify and leverage the transfe…
Action RecognitionDecoderFew-Shot action recognitionFew Shot Action Recognition+2