paper-with-me

홈 › Papers

Talk Less, Call Right: Enhancing Role-Play LLM Agents with Automatic Prompt Optimization and Role Prompting

2025-08-30 · Saksorn Ruangtanusak, Pittawat Taveekitworachai, Kunat Pipatanakul arxiv

This report investigates approaches for prompting a tool-augmented large language model (LLM) to act as a role-playing dialogue agent in the API track of the Commonsense Persona-grounded Dialogue Challenge (CPDC) 2025. In this setting, dialogue agents often produce overly long in-character responses (over-speaking) while failing to use tools effectively according to the persona (under-acting), such as generating function calls that do not exist or making unnecessary tool calls before answering. We explore four prompting approaches to address these issues: 1) basic role prompting, 2) improved role prompting, 3) automatic prompt optimization (APO), and 4) rule-based role prompting. The rule-based role prompting (RRP) approach achieved the best performance through two novel techniques-character-card/scene-contract design and strict enforcement of function calling-which led to an overall score of 0.571, improving on the zero-shot baseline score of 0.519. These findings demonstrate that RRP design can substantially improve the effectiveness and reliability of role-playing dialogue agents compared with more elaborate methods such as APO. To support future efforts in developing persona prompts, we are open-sourcing all of our best-performing prompts and the APO tool Source code is available at https://github.com/scb-10x/apo

📄 PDF Abstract BibTeX arXiv:2509.00482

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Enhanced Filterless Multi-Color VLC via QCT

2025-04-13 · Idris Cinemre, Serkan Vela, Gokce Hacioglu

Color shift keying (CSK) in visible light communication (VLC) often suffers from filter-induced crosstalk and reduced brightness. This paper proposes using quartered composite transform (QCT) with multi-color light-emitt…

R2-Talker: Realistic Real-Time Talking Head Synthesis with Hash Grid Landmarks Encoding and Progressive Multilayer Conditioning

2023-12-09 · Zhiling Ye, LiangGuo Zhang, Dingheng Zeng, Quan Lu 외

Dynamic NeRFs have recently garnered growing attention for 3D talking portrait synthesis. Despite advances in rendering speed and visual quality, challenges persist in enhancing efficiency and effectiveness. We present R…

Computational EfficiencyDiversityNeRF

DualTalk: Dual-Speaker Interaction for 3D Talking Head Conversations

2025-05-23 · CVPR 2025 1 · Ziqiao Peng, Yanbo Fan, HaoYu Wu, Xuan Wang 외

In face-to-face conversations, individuals need to switch between speaking and listening roles seamlessly. Existing 3D talking head generation models focus solely on speaking or listening, neglecting the natural dynamics…

Talking Head Generation

DiffTalker: Co-driven audio-image diffusion for talking faces via intermediate landmarks

2023-09-14 · Zipeng Qi, xulong Zhang, Ning Cheng, Jing Xiao 외

Generating realistic talking faces is a complex and widely discussed task with numerous applications. In this paper, we present DiffTalker, a novel model designed to generate lifelike talking faces through audio and land…

Face Generation

Enhancing Talk Moves Analysis in Mathematics Tutoring through Classroom Teaching Discourse

2024-12-18 · Jie Cao, Abhijit Suresh, Jennifer Jacobs, Charis Clevenger 외

Human tutoring interventions play a crucial role in supporting student learning, improving academic performance, and promoting personal growth. This paper focuses on analyzing mathematics tutoring discourse using talk mo…