paper-with-me

Papers

Do prompt positions really matter?

2023-05-23 · Junyu Mao, Stuart E. Middleton, Mahesan Niranjan

Prompt-based models have gathered a lot of attention from researchers due to their remarkable advancements in the fields of zero-shot and few-shot learning. Developing an effective prompt template plays a critical role. However, prior studies have mainly focused on prompt vocabulary searching or embedding initialization within a predefined template with the prompt position fixed. In this empirical study, we conduct the most comprehensive analysis to date of prompt position for diverse Natural Language Processing (NLP) tasks. Our findings quantify the substantial impact prompt position has on model performance. We observe that the prompt positions used in prior studies are often sub-optimal, and this observation is consistent even in widely used instruction-tuned models. These findings suggest prompt position optimisation as a valuable research direction to augment prompt engineering methodologies and prompt position-aware instruction tuning as a potential way to build more robust models in the future.

📄 PDF Abstract BibTeX arXiv:2305.14493

Code (1)

milliemaoo/prompt-position 공식 구현

Tasks

Few-Shot LearningNatural Language UnderstandingPositionPrompt Engineering

Similar Papers 제목 키워드 기반

Is Writing Prompts Really Making Art?

2023-01-26 · Jon McCormack, Camilo Cruz Gambardella, Nina Rajcic, Stephen James Krol 외

In recent years Generative Machine Learning systems have advanced significantly. A current wave of generative systems use text prompts to create complex imagery, video, even 3D datasets. The creators of these systems cla…

Position

Computational Models for Spatial Prepositions

2018-06-01 · WS 2018 6 · Georgiy Platonov, Lenhart Schubert

Developing computational models of spatial prepositions (such as on, in, above, etc.) is crucial for such tasks as human-machine collaboration, story understanding, and 3D model generation from descriptions. However, the…

When do Numbers Really Matter?

2014-08-07 · Hei Chan, Adnan Darwiche

Common wisdom has it that small distinctions in the probabilities quantifying a Bayesian network do not matter much for the resultsof probabilistic queries. However, one can easily develop realistic scenarios under which…

Does Feasibility Matter? Understanding the Impact of Feasibility on Synthetic Training Data

2025-05-15 · YiWen Liu, Jessica Bader, Jae Myung Kim

With the development of photorealistic diffusion models, models trained in part or fully on synthetic data achieve progressively better results. However, diffusion models still routinely generate images that would not ex…

AttributeLarge Language Model

On the test-time zero-shot generalization of vision-language models: Do we really need prompt learning?

2024-05-03 · CVPR 2024 1 · Maxime Zanella, Ismail Ben Ayed

The development of large vision-language models, notably CLIP, has catalyzed research into effective adaptation techniques, with a particular focus on soft prompt tuning. Conjointly, test-time augmentation, which utilize…

Computational EfficiencyPrompt LearningZero-shot Generalization