paper-with-me

홈 › Papers

POPI: Personalizing LLMs via Optimized Natural Language Preference Inference

2025-10-17 · Yizhuo Chen, Xin Liu, Ruijie Wang, Zheng Li, Pei Chen, Changlong Yu, Qingyu Yin, Priyanka Nigam, Meng Jiang, Bing Yin arxiv

Large language models (LLMs) are typically aligned with population-level preferences, despite substantial variation across individual users. We introduce POPI, a user-level personalization framework that separates the problem into two components connected by a natural-language interface: a shared inference model that distills heterogeneous user signals into a concise preference summary, and a shared generator that conditions on this summary to produce personalized responses. Both components are trained under a unified preference-optimization objective, with reinforcement learning handling the non-differentiable inference step. This objective decomposes into generator approximation error and summary informativeness, revealing how a single loss simultaneously drives accurate generation and informative summarization. Because the interface is natural language, learned summaries can be inferred once per user and reused across different generators -- including frozen, black-box commercial APIs. Across four personalization benchmarks, POPI generally improves personalization quality while reducing context overhead by up to an order of magnitude.

📄 PDF Abstract BibTeX arXiv:2510.17881

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

Laplace HypoPINN: Physics-Informed Neural Network for hypocenter localization and its predictive uncertainty

2022-05-28 · Muhammad Izzatullah, Isa Eren Yildirim, Umair bin Waheed, Tariq Alkhalifah

Several techniques have been proposed over the years for automatic hypocenter localization. While those techniques have pros and cons that trade-off computational efficiency and the susceptibility of getting trapped in l…

Computational Efficiency

Control of Legged Robots using Model Predictive Optimized Path Integral

2025-08-16 · Hossein Keshavarz, Alejandro Ramirez-Serrano, Majid Khadiv arxiv

Legged robots possess a unique ability to traverse rough terrains and navigate cluttered environments, making them well-suited for complex, real-world unstructured scenarios. However, such robots have not yet achieved th…

TopoPilot: Reliable Conversational Workflow Automation for Topological Data Analysis and Visualization

2026-03-26 · Nathaniel Gorski, Shusen Liu, Bei Wang arxiv

Recent agentic systems demonstrate that large language models can generate scientific visualizations from natural language. However, reliability remains a major limitation: systems may execute invalid operations, introdu…

Personalizing Multimodal Large Language Models for Image Captioning: An Experimental Analysis

2024-12-04 · Davide Bucciarelli, Nicholas Moratelli, Marcella Cornia, Lorenzo Baraldi 외

The task of image captioning demands an algorithm to generate natural language descriptions of visual inputs. Recent advancements have seen a convergence between image captioning research and the development of Large Lan…

Image CaptioningImage DescriptionPrompt Learning

CoPL: Collaborative Preference Learning for Personalizing LLMs

2025-03-03 · Youngbin Choi, Seunghyuk Cho, Minjong Lee, Moonjeong Park 외

Personalizing large language models (LLMs) is important for aligning outputs with diverse user preferences, yet existing methods struggle with flexibility and generalization. We propose CoPL (Collaborative Preference Lea…

Collaborative Filtering