paper-with-me

홈 › Papers

ViPer: Visual Personalization of Generative Models via Individual Preference Learning

2024-07-24 · Sogand Salehi, Mahdi Shafiei, Teresa Yeo, Roman Bachmann, Amir Zamir

Different users find different images generated for the same prompt desirable. This gives rise to personalized image generation which involves creating images aligned with an individual's visual preference. Current generative models are, however, unpersonalized, as they are tuned to produce outputs that appeal to a broad audience. Using them to generate images aligned with individual users relies on iterative manual prompt engineering by the user which is inefficient and undesirable. We propose to personalize the image generation process by first capturing the generic preferences of the user in a one-time process by inviting them to comment on a small selection of images, explaining why they like or dislike each. Based on these comments, we infer a user's structured liked and disliked visual attributes, i.e., their visual preference, using a large language model. These attributes are used to guide a text-to-image model toward producing images that are tuned towards the individual user's visual preference. Through a series of user studies and large language model guided evaluations, we demonstrate that the proposed method results in generations that are well aligned with individual users' visual preferences.

📄 PDF Abstract BibTeX arXiv:2407.17365

Code (0)

등록된 구현이 없습니다.

Tasks

Image GenerationLanguage ModelingLanguage ModellingLarge Language ModelPersonalized Image GenerationPrompt Engineering

Similar Papers 제목 키워드 기반

Premier: Personalized Preference Modulation with Learnable User Embedding in Text-to-Image Generation

2026-03-21 · Zihao Wang, Yuxiang Wei, Xinpeng Zhou, Tianyu Zhang 외 arxiv

Text-to-image generation has advanced rapidly, yet it still struggles to capture the nuanced user preferences. Existing approaches typically rely on multimodal large language models to infer user preferences, but the der…

Personalized Image GenerationText-to-Image Generation

ImageGem: In-the-wild Generative Image Interaction Dataset for Generative Model Personalization

2025-10-21 · Yuanhe Guo, Linxi Xie, Zhuoran Chen, Kangrui Yu 외 arxiv

We introduce ImageGem, a dataset for studying generative models that understand fine-grained individual preferences. We posit that a key challenge hindering the development of such a generative model is the lack of in-th…

Image Retrieval

DesignPref: Capturing Personal Preferences in Visual Design Generation

2025-11-25 · Yi-Hao Peng, Jeffrey P. Bigham, Jason Wu arxiv

Generative models, such as large language models and text-to-image diffusion models, are increasingly used to create visual designs like user interfaces (UIs) and presentation slides. Finetuning and benchmarking these ge…

VIPER: Visual In-Context Physics Reasoning for Physically Plausible Video Generation

2026-07-26 · Tianxiao Chen, Hanmo Chen, Huajin Chen, Bo Li 외 arxiv

Modern video generation models can synthesize visually compelling and temporally coherent clips, yet controlling their physical behavior remains difficult with standard text and image conditions. The core challenge is a …

Video Generation

Efficient Personalization of Generative User Interfaces

2026-04-10 · Yi-Hao Peng, Samarth Das, Jeffrey P. Bigham, Jason Wu arxiv

Generative user interfaces (UIs) create new opportunities to adapt interfaces to individual users on demand, but personalization remains difficult because desirable UI properties are subjective, hard to articulate, and c…