Divergent Search for Few-Shot Image Classification
When data is unlabelled and the target task is not known a priori, divergent search offers a strategy for learning a wide range of skills. Having such a repertoire allows a system to adapt to new, unforeseen tasks. Unlabelled image data is plentiful, but it is not always known which features will be required for downstream tasks. We propose a method for divergent search in the few-shot image classification setting and evaluate with Omniglot and Mini-ImageNet. This high-dimensional behavior space includes all possible ways of partitioning the data. To manage divergent search in this space, we rely on a meta-learning framework to integrate useful features from diverse tasks into a single model. The final layer of this model is used as an index into the `archive' of all past behaviors. We search for regions in the behavior space that the current archive cannot reach. As expected, divergent search is outperformed by models with a strong bias toward the evaluation tasks. But it is able to match and sometimes exceed the performance of models that have a weak bias toward the target task or none at all. This demonstrates that divergent search is a viable approach, even in high-dimensional behavior spaces.
Code (0)
등록된 구현이 없습니다.
Tasks
ClassificationFew-Shot Image ClassificationGeneral Classificationimage-classificationImage ClassificationMeta-LearningSimilar Papers 제목 키워드 기반
LDRE: LLM-based Divergent Reasoning and Ensemble for Zero-Shot Composed Image Retrieval
Zero-Shot Composed Image Retrieval (ZS-CIR) has garnered increasing interest in recent years, which aims to retrieve a target image based on a query composed of a reference image and a modification text without training …
Image RetrievalImage to textRetrievalZero-Shot Composed Image Retrieval (ZS-CIR)From Shots to Stories: LLM-Assisted Video Editing with Unified Language Representations
Large Language Models (LLMs) and Vision-Language Models (VLMs) have demonstrated remarkable reasoning and generalization capabilities in video understanding; however, their application in video editing remains largely un…
Video EditingVideo UnderstandingDANet: Divergent Activation for Weakly Supervised Object Localization
Weakly supervised object localization remains a challenge when learning object localization models from image category labels. Optimizing image classification tends to activate object parts and ignore the full object ext…
ClassificationGeneral Classificationimage-classificationImage Classification+3DiCoRe: Enhancing Zero-shot Event Detection via Divergent-Convergent LLM Reasoning
Zero-shot Event Detection (ED), the task of identifying event mentions in natural language text without any training data, is critical for document understanding in specialized domains. Understanding the complex event on…
document understandingEvent DetectionTransfer LearningRetrieval-style In-Context Learning for Few-shot Hierarchical Text Classification
Hierarchical text classification (HTC) is an important task with broad applications, while few-shot HTC has gained increasing interest recently. While in-context learning (ICL) with large language models (LLMs) has achie…
Contrastive Learningfew-shot-htcFew-shot HTCFew-Shot Learning+7