paper-with-me

홈 › Papers

Input-Tuning: Adapting Unfamiliar Inputs to Frozen Pretrained Models

2022-03-07 · Shengnan An, Yifei Li, Zeqi Lin, Qian Liu, Bei Chen, Qiang Fu, Weizhu Chen, Nanning Zheng, Jian-Guang Lou

Recently the prompt-tuning paradigm has attracted significant attention. By only tuning continuous prompts with a frozen pre-trained language model (PLM), prompt-tuning takes a step towards deploying a shared frozen PLM to serve numerous downstream tasks. Although prompt-tuning shows good performance on certain natural language understanding (NLU) tasks, its effectiveness on natural language generation (NLG) tasks is still under-explored. In this paper, we argue that one of the factors hindering the development of prompt-tuning on NLG tasks is the unfamiliar inputs (i.e., inputs are linguistically different from the pretraining corpus). For example, our preliminary exploration reveals a large performance gap between prompt-tuning and fine-tuning when unfamiliar inputs occur frequently in NLG tasks. This motivates us to propose input-tuning, which fine-tunes both the continuous prompts and the input representations, leading to a more effective way to adapt unfamiliar inputs to frozen PLMs. Our proposed input-tuning is conceptually simple and empirically powerful. Experimental results on seven NLG tasks demonstrate that input-tuning is significantly and consistently better than prompt-tuning. Furthermore, on three of these tasks, input-tuning can achieve a comparable or even better performance than fine-tuning.

📄 PDF Abstract BibTeX arXiv:2203.03131

Code (0)

등록된 구현이 없습니다.

Tasks

Language ModelingLanguage ModellingNatural Language UnderstandingText Generation

Similar Papers 제목 키워드 기반

Prompt Generation Networks for Input-Space Adaptation of Frozen Vision Transformers

2022-10-12 · Jochem Loedeman, Maarten C. Stol, Tengda Han, Yuki M. Asano

With the introduction of the transformer architecture in computer vision, increasing model scale has been demonstrated as a clear path to achieving performance and robustness gains. However, with model parameter counts r…

Prompt LearningTransfer Learning

Knowledge-Preserved Model Tuning in Null-Space for Robust Spatio-Temporal Video Grounding

2026-06-02 · Haoxuan Chen, Xianqin Liu, Jian-Fang Hu arxiv

Spatio-Temporal Video Grounding aims to localize object tubes based on textual queries. While recent methods have achieved remarkable success, they mainly focus on high-quality(HQ) inputs, neglecting the widespread prese…

Spatio-Temporal Video Grounding

MAGPrompt: Message-Adaptive Graph Prompt Tuning for Graph Neural Networks

2026-02-05 · Long D. Nguyen, Binh P. Nguyen arxiv

Pre-trained graph neural networks (GNNs) transfer well, but adapting them to downstream tasks remains challenging due to mismatches between pre-training objectives and task requirements. Graph prompt tuning offers a para…

DINO-MVR: Multi-View Readout of Frozen DINOv3 for Annotation-Efficient Medical Segmentation

2026-05-08 · Wei Jiang, Feng Liu, Nan Ye, Hongfu Sun arxiv

Adapting foundation models to medical segmentation typically requires either backbone fine-tuning or high-capacity task-specific decoders, both of which are difficult to fit reliably when annotations are scarce. We show …

Tumor Segmentation

Chain of Operators: An Inference-Time Harness for In-Context Operator Learning

2026-06-10 · Minghui Yang, Ling Guo, Liu Yang arxiv

While scientific foundation models show immense promise in accelerating physical simulations and numerical forecasting, they remain notoriously brittle when encountering out-of-distribution (OOD) scenarios. Adapting thes…