paper-with-me

Papers

HyPerAlign: Interpretable Personalized LLM Alignment via Hypothesis Generation

2025-04-29 · Cristina Garbacea, Chenhao Tan

Alignment algorithms are widely used to align large language models (LLMs) to human users based on preference annotations. Typically these (often divergent) preferences are aggregated over a diverse set of users, resulting in fine-tuned models that are aligned to the ``average-user'' preference. Nevertheless, current models are used by individual users in very specific contexts and situations, emphasizing the need for user-dependent preference control. In this work we address the problem of personalizing LLM outputs to their users. We aim to generate customized responses tailored to specific individuals instead of generic outputs that emulate the collective voices of diverse populations. We propose HyPerAlign, an interpretable and sample-efficient hypothesis-driven personalization approach for LLM models. Given few-shot examples written by a particular user, we first infer hypotheses about their communication strategies, personality, and writing style, then prompt LLM models with these hypotheses and user-specific attributes to generate customized outputs. We conduct experiments on two different personalization tasks, namely authorship attribution and deliberative alignment, with datasets from diverse domains (news articles, blog posts, emails, jailbreaking benchmarks). Results demonstrate the superiority of hypothesis-driven LLM personalization compared to preference-based fine-tuning methods. For authorship attribution, HyPerAlign generations have consistently high win-rates (commonly $> 90\%$) against state-of-the-art preference fine-tuning approaches across diverse user profiles and LLM models. For deliberative alignment, the helpfulness of LLM models is improved by up to $70\%$ on average. Overall, HyPerAlign represents an interpretable and sample-efficient strategy for the personalization of LLM models to individual users.

📄 PDF Abstract BibTeX arXiv:2505.00038

Code (0)

등록된 구현이 없습니다.

Tasks

ArticlesAuthorship Attribution

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically
ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

Kernel Hyperalignment

2012-12-01 · NeurIPS 2012 12 · Alexander Lorbert, Peter J. Ramadge

We offer a regularized, kernel extension of the multi-set, orthogonal Procrustes problem, or hyperalignment. Our new method, called Kernel Hyperalignment, expands the scope of hyperalignment to include nonlinear measures…

HyperAlign: Hypernetwork for Efficient Test-Time Alignment of Diffusion Models

2026-01-22 · Xin Xie, Jiaxian Guo, Dong Gong arxiv

Diffusion model alignment aims to bridge the gap between generated outputs and human preferences by enhancing both semantic consistency with textual prompts and overall visual quality. Existing alignment methods face a c…

Computational Efficiency

HyperAlign: Hyperbolic Entailment Cones for Adaptive Text-to-Image Alignment Assessment

2026-01-08 · Wenzhi Chen, Bo Hu, Leida Li, Lihuo He 외 arxiv

With the rapid development of text-to-image generation technology, accurately assessing the alignment between generated images and text prompts has become a critical challenge. Existing methods rely on Euclidean space me…

Text-to-Image Generation

Deep Hyperalignment

2017-10-11 · NeurIPS 2017 12 · Muhammad Yousefnezhad, Daoqiang Zhang

This paper proposes Deep Hyperalignment (DHA) as a regularized, deep extension, scalable Hyperalignment (HA) method, which is well-suited for applying functional alignment to fMRI datasets with nonlinearity, high-dimensi…

Supervised Hyperalignment for multi-subject fMRI data alignment

2020-01-09 · Muhammad Yousefnezhad, Alessandro Selvitella, Liangxiu Han, Daoqiang Zhang

Hyperalignment has been widely employed in Multivariate Pattern (MVP) analysis to discover the cognitive states in the human brains based on multi-subject functional Magnetic Resonance Imaging (fMRI) datasets. Most of th…

Multi-Subject Fmri Data AlignmentTime SeriesTime Series Analysis