paper-with-me

홈 › Papers

GRL-Prompt: Towards Knowledge Graph based Prompt Optimization via Reinforcement Learning

2024-11-19 · Yuze Liu, Tingjie Liu, Tiehua Zhang, Youhua Xia, Jinze Wang, Zhishu Shen, Jiong Jin, Fei Richard Yu

Large language models (LLMs) have demonstrated impressive success in a wide range of natural language processing (NLP) tasks due to their extensive general knowledge of the world. Recent works discovered that the performance of LLMs is heavily dependent on the input prompt. However, prompt engineering is usually done manually in a trial-and-error fashion, which can be labor-intensive and challenging in order to find the optimal prompts. To address these problems and unleash the utmost potential of LLMs, we propose a novel LLMs-agnostic framework for prompt optimization, namely GRL-Prompt, which aims to automatically construct optimal prompts via reinforcement learning (RL) in an end-to-end manner. To provide structured action/state representation for optimizing prompts, we construct a knowledge graph (KG) that better encodes the correlation between the user query and candidate in-context examples. Furthermore, a policy network is formulated to generate the optimal action by selecting a set of in-context examples in a rewardable order to construct the prompt. Additionally, the embedding-based reward shaping is utilized to stabilize the RL training process. The experimental results show that GRL-Prompt outperforms recent state-of-the-art methods, achieving an average increase of 0.10 in ROUGE-1, 0.07 in ROUGE-2, 0.07 in ROUGE-L, and 0.05 in BLEU.

📄 PDF Abstract BibTeX arXiv:2411.14479

Code (0)

등록된 구현이 없습니다.

Tasks

General KnowledgePrompt EngineeringReinforcement Learning (RL)

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

RELIEF: Reinforcement Learning Empowered Graph Feature Prompt Tuning

2024-08-06 · Jiapeng Zhu, Zichen Ding, Jianxiang Yu, Jiaqi Tan 외

The advent of the "pre-train, prompt" paradigm has recently extended its generalization ability and data efficiency to graph representation learning, following its achievements in Natural Language Processing (NLP). Initi…

Combinatorial OptimizationGraph Neural NetworkGraph Representation Learningreinforcement-learning+4

Plug-and-Play PPO: An Adaptive Point Prompt Optimizer Making SAM Greater

2025-01-01 · CVPR 2025 1 · Xueyu Liu, Rui Wang, Yexin Lai, Guangze Shi 외

Powered by extensive curated training data, the Segment Anything Model (SAM) demonstrates impressive generalization capabilities in open-world scenarios, effectively guided by user-provided prompts. However, the clas…

Deep Reinforcement LearningSegmentation

MultiPrompter: Cooperative Prompt Optimization with Multi-Agent Reinforcement Learning

2023-10-25 · Dong-Ki Kim, Sungryull Sohn, Lajanugen Logeswaran, Dongsub Shim 외

Recently, there has been an increasing interest in automated prompt optimization based on reinforcement learning (RL). This approach offers important advantages, such as generating interpretable prompts and being compati…

Multi-agent Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

TKG-Thinker: Towards Dynamic Reasoning over Temporal Knowledge Graphs via Agentic Reinforcement Learning

2026-02-05 · Zihao Jiang, Miao Peng, Zhenyan Shan, Wenjie Xu 외 arxiv

Temporal knowledge graph question answering (TKGQA) aims to answer time-sensitive questions by leveraging temporal knowledge bases. While Large Language Models (LLMs) demonstrate significant potential in TKGQA, current p…

Graph Question AnsweringReinforcement LearningKnowledge Graphs

Agentic-KGR: Co-evolutionary Knowledge Graph Construction through Multi-Agent Reinforcement Learning

2025-10-10 · Jing Li, Zhijie Sun, Zhicheng Zhou, Suming Qiu 외 arxiv

Current knowledge-enhanced large language models (LLMs) rely on static, pre-constructed knowledge bases that suffer from coverage gaps and temporal obsolescence, limiting their effectiveness in dynamic information enviro…

Multi-agent Reinforcement LearningKnowledge Graphs