paper-with-me

홈 › Papers

Whose Boat Does it Float? Improving Personalization in Preference Tuning via Inferred User Personas

2025-01-20 · Nishant Balepur, Vishakh Padmakumar, Fumeng Yang, Shi Feng, Rachel Rudinger, Jordan Lee Boyd-Graber

LLMs are tuned to follow instructions (aligned) by learning which of two outputs users prefer for a prompt. However, this preference data format does not convey why users prefer responses that are chosen or rejected, so LLMs trained on these datasets cannot tailor responses to varied user needs. To surface these parameters of personalization, we apply abductive reasoning to preference data, inferring needs and interests of users, i.e. personas, that may prefer each output. We test this idea in two steps: Persona Inference (PI)-abductively inferring personas of users who prefer chosen or rejected outputs-and Persona Tailoring (PT)-training models to tailor responses to personas from PI. We find: 1) LLMs infer personas accurately explaining why different users may prefer both chosen or rejected outputs; 2) Training on preference data augmented with PI personas via PT boosts personalization, enabling models to support user-written personas; and 3) Rejected response personas form harder personalization evaluations, showing PT better aids users with uncommon preferences versus typical alignment methods. We argue for an abductive view of preferences for personalization, asking not only which response is better but when, why, and for whom.

📄 PDF Abstract BibTeX arXiv:2501.11549

Code (1)

pinafore/alignment-personalization 공식 구현

Similar Papers 제목 키워드 기반

HyperTrace: Hypothesis-Based Preference Tracing for Online LLM Personalization

2026-09-09 · Jianzhi Shen, Keyu Mao, Minghao Shao, Chuanyang Jin 외 arxiv

Personalized language models aim to adapt responses to individual users, whose preferences are often latent and revealed gradually through interaction. Existing training-free methods rely on stored histories or retrieved…

TurboAttention: Efficient Attention Approximation For High Throughputs LLMs

2024-12-11 · Hao Kang, Srikant Bharadwaj, James Hensman, Tushar Krishna 외

Large language model (LLM) inference demands significant amount of computation and memory, especially in the key attention mechanism. While techniques, such as quantization and acceleration algorithms, like FlashAttentio…

Computational EfficiencyLanguage ModelingLanguage ModellingLarge Language Model+1

Bridging Personalization and Control in Scientific Personalized Search

2024-11-05 · Sheshera Mysore, Garima Dhanania, Kishor Patil, Surya Kallumadi 외

Personalized search is a problem where models benefit from learning user preferences from per-user historical interaction data. The inferred preferences enable personalized ranking models to improve the relevance of docu…

Retrieval

Preference Heads in Large Language Models: A Mechanistic Framework for Interpretable Personalization

2026-04-24 · Weixu Zhang, Ye Yuan, Changjiang Han, Yuxing Tian 외 arxiv

Large Language Models (LLMs) exhibit strong implicit personalization ability, yet most existing approaches treat this behavior as a black box, relying on prompt engineering or fine tuning on user data. In this work, we a…

Prompt Engineering

Towards Effective Model Editing for LLM Personalization

2025-12-15 · Baixiang Huang, Limeng Cui, Jiapeng Liu, Haoran Wang 외 arxiv

Personalization is becoming indispensable for LLMs to align with individual user preferences and needs. Yet current approaches are often computationally expensive, data-intensive, susceptible to catastrophic forgetting, …

Computational EfficiencyQuestion Answering