paper-with-me

홈 › Papers

Prompt Waywardness: The Curious Case of Discretized Interpretation of Continuous Prompts

2021-12-15 · NAACL 2022 7 · Daniel Khashabi, Shane Lyu, Sewon Min, Lianhui Qin, Kyle Richardson, Sean Welleck, Hannaneh Hajishirzi, Tushar Khot, Ashish Sabharwal, Sameer Singh, Yejin Choi

Fine-tuning continuous prompts for target tasks has recently emerged as a compact alternative to full model fine-tuning. Motivated by these promising results, we investigate the feasibility of extracting a discrete (textual) interpretation of continuous prompts that is faithful to the problem they solve. In practice, we observe a "wayward" behavior between the task solved by continuous prompts and their nearest neighbor discrete projections: We can find continuous prompts that solve a task while being projected to an arbitrary text (e.g., definition of a different or even a contradictory task), while being within a very small (2%) margin of the best continuous prompt of the same size for the task. We provide intuitions behind this odd and surprising behavior, as well as extensive empirical analyses quantifying the effect of various parameters. For instance, for larger model sizes we observe higher waywardness, i.e, we can find prompts that more closely map to any arbitrary text with a smaller drop in accuracy. These findings have important implications relating to the difficulty of faithfully interpreting continuous prompts and their generalization across models and tasks, providing guidance for future progress in prompting language models.

📄 PDF Abstract BibTeX arXiv:2112.08348

Code (1)

alrope123/prompt-waywardness 공식 구현 pytorch

Similar Papers 제목 키워드 기반

PROMPT WAYWARDNESS: The Curious Case of Discretized Interpretation of Continuous Prompts

2022-01-16 · ACL ARR January 2022 1 · Anonymous

Fine-tuning continuous prompts for target tasks has recently emerged as a compact alternative to full model fine-tuning. Motivated by these promising results, we investigate the feasibility of extracting a discrete (text…

Exploring the Curious Case of Code Prompts

2023-04-26 · Li Zhang, Liam Dugan, Hainiu Xu, Chris Callison-Burch

Recent work has shown that prompting language models with code-like representations of natural language leads to performance improvements on structured reasoning tasks. However, such tasks comprise only a small subset of…

The Impact of Using Regression Models to Build Defect Classifiers

2022-02-12 · Gopi Krishnan Rajbahadur, Shaowei Wang, Yasutaka Kamei, Ahmed E. Hassan

It is common practice to discretize continuous defect counts into defective and non-defective classes and use them as a target variable when building defect classifiers (discretized classifiers). However, this discretiza…

regression

The Curious Case of Metonymic Verbs: A Distributional Characterization

2013-03-01 · WS 2013 3 · Jason Utt, Aless Lenci, ro, Sebastian Pad{\'o} 외

Word Embeddings vs Word Types for Sequence Labeling: the Curious Case of CV Parsing

2015-06-01 · WS 2015 6 · Melanie Tosik, Carsten Lygteskov Hansen, Gerard Goossen, Mihai Rotaru
Word Embeddings