paper-with-me

홈 › Papers

Helpful assistant or fruitful facilitator? Investigating how personas affect language model behavior

2024-07-02 · Pedro Henrique Luz de Araujo, Benjamin Roth

One way to personalize and steer generations from large language models (LLM) is to assign a persona: a role that describes how the user expects the LLM to behave (e.g., a helpful assistant, a teacher, a woman). This paper investigates how personas affect diverse aspects of model behavior. We assign to seven LLMs 162 personas from 12 categories spanning variables like gender, sexual orientation, and occupation. We prompt them to answer questions from five datasets covering objective (e.g., questions about math and history) and subjective tasks (e.g., questions about beliefs and values). We also compare persona's generations to two baseline settings: a control persona setting with 30 paraphrases of "a helpful assistant" to control for models' prompt sensitivity, and an empty persona setting where no persona is assigned. We find that for all models and datasets, personas show greater variability than the control setting and that some measures of persona behavior generalize across models.

📄 PDF Abstract BibTeX arXiv:2407.02099

Code (1)

peluz/persona-behavior 공식 구현

Tasks

Language ModelingLanguage ModellingMath

Similar Papers 제목 키워드 기반

When "A Helpful Assistant" Is Not Really Helpful: Personas in System Prompts Do Not Improve Performances of Large Language Models

2023-11-16 · Mingqian Zheng, Jiaxin Pei, Lajanugen Logeswaran, Moontae Lee 외

Prompting serves as the major way humans interact with Large Language Models (LLM). Commercial AI systems commonly define the role of the LLM in system prompts. For example, ChatGPT uses ``You are a helpful assistant'' a…

The Assistant Axis: Situating and Stabilizing the Default Persona of Language Models

2026-01-15 · Christina Lu, Jack Gallagher, Jonathan Michala, Kyle Fish 외 arxiv

Large language models can represent a variety of personas but typically default to a helpful Assistant identity cultivated during post-training. We investigate the structure of the space of model personas by extracting a…

Probing Persona-Dependent Preferences in Language Models

2026-05-13 · Oscar Gilg, Pierre Beckmann, Daniel Paleka, Patrick Butlin arxiv

Large language models (LLMs) can be said to have preferences: they reliably pick certain tasks and outputs over others, and preferences shaped by post-training and system prompts appear to shape much of their behaviour. …

"Many Are My Names": The Anatomy of the Assistant and Its Personas via Sparse Autoencoders

2026-08-08 · Adelaide Danilov, Aria Nourbakhsh, Oleksandr Marchenko Breneur, Salima Lamsiyah arxiv

How a language model internally represents who is speaking, the Assistant, an assigned roleplay persona, or a narrated story character, remains underexplored. We study speaker representations using a dataset of user-expr…

Humanlike Multi-user Agent (HUMA): Designing a Deceptively Human AI Facilitator for Group Chats

2025-11-21 · Mateusz Jacniacki, Martí Carmona Serrat arxiv

Conversational agents built on large language models (LLMs) are becoming increasingly prevalent, yet most systems are designed for one-on-one, turn-based exchanges rather than natural, asynchronous group chats. As AI ass…