paper-with-me

홈 › Papers

Concentrate Attention: Towards Domain-Generalizable Prompt Optimization for Language Models

2024-06-15 · Chengzhengxu Li, Xiaoming Liu, Zhaohan Zhang, Yichen Wang, Chen Liu, Yu Lan, Chao Shen

Recent advances in prompt optimization have notably enhanced the performance of pre-trained language models (PLMs) on downstream tasks. However, the potential of optimized prompts on domain generalization has been under-explored. To explore the nature of prompt generalization on unknown domains, we conduct pilot experiments and find that (i) Prompts gaining more attention weight from PLMs' deep layers are more generalizable and (ii) Prompts with more stable attention distributions in PLMs' deep layers are more generalizable. Thus, we offer a fresh objective towards domain-generalizable prompts optimization named "Concentration", which represents the "lookback" attention from the current decoding token to the prompt tokens, to increase the attention strength on prompts and reduce the fluctuation of attention distribution. We adapt this new objective to popular soft prompt and hard prompt optimization methods, respectively. Extensive experiments demonstrate that our idea improves comparison prompt optimization methods by 1.42% for soft prompt generalization and 2.16% for hard prompt generalization in accuracy on the multi-source domain generalization setting, while maintaining satisfying in-domain performance. The promising results validate the effectiveness of our proposed prompt optimization objective and provide key insights into domain-generalizable prompts.

📄 PDF Abstract BibTeX arXiv:2406.10584

Code (1)

czx-li/Concentrate-Attention 공식 구현 pytorch

Tasks

Domain Generalization

Similar Papers 제목 키워드 기반

Generalizable Engagement Estimation in Conversation via Domain Prompting and Parallel Attention

2025-08-20 · Yangche Yu, Yin Chen, Jia Li, Peng Jia 외 arxiv

Accurate engagement estimation is essential for adaptive human-computer interaction systems, yet robust deployment is hindered by poor generalizability across diverse domains and challenges in modeling complex interactio…

Localist LLMs -- A Mathematical Framework for Dynamic Locality Control

2025-10-10 · Joachim Diederich arxiv

We present a novel framework for training large language models with continuously adjustable internal representations that span the full spectrum from localist (interpretable, rule-based) to distributed (generalizable, e…

CFPL-FAS: Class Free Prompt Learning for Generalizable Face Anti-spoofing

2024-03-21 · CVPR 2024 1 · Ajian Liu, Shuai Xue, Jianwen Gan, Jun Wan 외

Domain generalization (DG) based Face Anti-Spoofing (FAS) aims to improve the model's performance on unseen domains. Existing methods either rely on domain labels to align domain-invariant feature spaces, or disentangle …

Domain GeneralizationFace Anti-SpoofingPrompt Learning

Prompting Segment Anything Model with Domain-Adaptive Prototype for Generalizable Medical Image Segmentation

2024-09-19 · Zhikai Wei, Wenhui Dong, Peilin Zhou, Yuliang Gu 외

Deep learning based methods often suffer from performance degradation caused by domain shift. In recent years, many sophisticated network structures have been designed to tackle this problem. However, the advent of large…

Domain GeneralizationImage SegmentationMedical Image SegmentationSegmentation+3

Generalizable Self-Evolving Memory for Automatic Prompt Optimization

2026-03-23 · Guanbao Liang, Yuanchen Bei, Sheng Zhou, Yuheng Qin 외 arxiv

Automatic prompt optimization is a promising approach for adapting large language models (LLMs) to downstream tasks, yet existing methods typically search for a specific prompt specialized to a fixed task. This paradigm …