ProtoCon: Pseudo-label Refinement via Online Clustering and Prototypical Consistency for Efficient Semi-supervised Learning
Confidence-based pseudo-labeling is among the dominant approaches in semi-supervised learning (SSL). It relies on including high-confidence predictions made on unlabeled data as additional targets to train the model. We propose ProtoCon, a novel SSL method aimed at the less-explored label-scarce SSL where such methods usually underperform. ProtoCon refines the pseudo-labels by leveraging their nearest neighbours' information. The neighbours are identified as the training proceeds using an online clustering approach operating in an embedding space trained via a prototypical loss to encourage well-formed clusters. The online nature of ProtoCon allows it to utilise the label history of the entire dataset in one training cycle to refine labels in the following cycle without the need to store image embeddings. Hence, it can seamlessly scale to larger datasets at a low cost. Finally, ProtoCon addresses the poor training signal in the initial phase of training (due to fewer confident predictions) by introducing an auxiliary self-supervised loss. It delivers significant gains and faster convergence over state-of-the-art across 5 datasets, including CIFARs, ImageNet and DomainNet.
Code (0)
등록된 구현이 없습니다.
Tasks
Online ClusteringPseudo LabelSimilar Papers 제목 키워드 기반
ProtoConNet: Prototypical Augmentation and Alignment for Open-Set Few-Shot Image Classification
Open-set few-shot image classification aims to train models using a small amount of labeled data, enabling them to achieve good generalization when confronted with unknown environments. Existing methods mainly use visual…
Few-Shot Image ClassificationRepresentation LearningSelf-supervised Reflective Learning through Self-distillation and Online Clustering for Speaker Representation Learning
Speaker representation learning is crucial for voice recognition systems, with recent advances in self-supervised approaches reducing dependency on labeled data. Current two-stage iterative frameworks, while effective, s…
ClusteringKnowledge DistillationOnline ClusteringPseudo Label+1Neighbour Consistency Guided Pseudo-Label Refinement for Unsupervised Person Re-Identification
Unsupervised person re-identification (ReID) aims at learning discriminative identity features for person retrieval without any annotations. Recent advances accomplish this task by leveraging clustering-based pseudo labe…
ClusteringPerson Re-IdentificationPerson RetrievalPseudo Label+2Unsupervised Domain Adaptation with Dynamic Clustering and Contrastive Refinement for Gait Recognition
Gait recognition is an emerging identification technology that distinguishes individuals at long distances by analyzing individual walking patterns. Traditional techniques rely heavily on large-scale labeled datasets, wh…
ClusteringDomain AdaptationGait RecognitionPseudo Label+1Pseudo-label Refinement for Improving Self-Supervised Learning Systems
Self-supervised learning systems have gained significant attention in recent years by leveraging clustering-based pseudo-labels to provide supervision without the need for human annotations. However, the noise in these p…
ClusteringDomain AdaptationPerson Re-IdentificationPseudo Label+2