paper-with-me

홈 › Papers

What Makes Good Few-shot Examples for Vision-Language Models?

2024-05-22 · Zhaojun Guo, Jinghui Lu, Xuejing Liu, Rui Zhao, Zhenxing Qian, Fei Tan

Despite the notable advancements achieved by leveraging pre-trained vision-language (VL) models through few-shot tuning for downstream tasks, our detailed empirical study highlights a significant dependence of few-shot learning outcomes on the careful selection of training examples - a facet that has been previously overlooked in research. In this study, we delve into devising more effective strategies for the meticulous selection of few-shot training examples, as opposed to relying on random sampling, to enhance the potential of existing few-shot prompt learning methodologies. To achieve this, we assess the effectiveness of various Active Learning (AL) techniques for instance selection, such as Entropy and Margin of Confidence, within the context of few-shot training. Furthermore, we introduce two innovative selection methods - Representativeness (REPRE) and Gaussian Monte Carlo (Montecarlo) - designed to proactively pinpoint informative examples for labeling in relation to pre-trained VL models. Our findings demonstrate that both REPRE and Montecarlo significantly surpass both random selection and AL-based strategies in few-shot training scenarios. The research also underscores that these instance selection methods are model-agnostic, offering a versatile enhancement to a wide array of few-shot training methodologies.

📄 PDF Abstract BibTeX arXiv:2405.13532

Code (0)

등록된 구현이 없습니다.

Tasks

Active LearningFew-Shot LearningPrompt Learning

Similar Papers 제목 키워드 기반

What Makes Good Examples for Visual In-Context Learning?

2023-01-31 · NeurIPS 2023 11 · Yuanhan Zhang, Kaiyang Zhou, Ziwei Liu

Large-scale models trained on broad data have recently become the mainstream architecture in computer vision due to their strong generalization performance. In this paper, the main focus is on an emergent ability in larg…

In-Context LearningRetrieval

What Makes Good In-Context Examples for GPT-$3$?

2021-01-17 · Jiachang Liu, Dinghan Shen, Yizhe Zhang, Bill Dolan 외

GPT-$3$ has attracted lots of attention due to its superior performance across a wide range of NLP tasks, especially with its powerful and versatile in-context few-shot learning ability. Despite its success, we found tha…

Few-Shot LearningNatural Language UnderstandingOpen-Domain Question AnsweringQuestion Answering+4

What Makes for Good Visual Instructions? Synthesizing Complex Visual Reasoning Instructions for Visual Instruction Tuning

2023-11-02 · Yifan Du, Hangyu Guo, Kun Zhou, Wayne Xin Zhao 외

Visual instruction tuning is an essential approach to improving the zero-shot generalization capability of Multi-modal Large Language Models (MLLMs). A surge of visual instruction datasets with various focuses and charac…

MMEVisual ReasoningZero-shot Generalization

What Makes Language Models Good-enough?

2024-06-06 · Daiki Asami, Saku Sugawara

Psycholinguistic research suggests that humans may build a representation of linguistic input that is 'good-enough' for the task at hand. This study examines what architectural features make language models learn human-l…

What makes ImageNet good for transfer learning?

2016-08-30 · Minyoung Huh, Pulkit Agrawal, Alexei A. Efros

The tremendous success of ImageNet-trained deep features on a wide range of transfer tasks begs the question: what are the properties of the ImageNet dataset that are critical for learning good, general-purpose features?…

Action ClassificationGeneral ClassificationScene ClassificationTransfer Learning