paper-with-me

Papers

ImageGem: In-the-wild Generative Image Interaction Dataset for Generative Model Personalization

2025-10-21 · Yuanhe Guo, Linxi Xie, Zhuoran Chen, Kangrui Yu, Ryan Po, Guandao Yang, Gordon Wetztein, Hongyi Wen arxiv

We introduce ImageGem, a dataset for studying generative models that understand fine-grained individual preferences. We posit that a key challenge hindering the development of such a generative model is the lack of in-the-wild and fine-grained user preference annotations. Our dataset features real-world interaction data from 57K users, who collectively have built 242K customized LoRAs, written 3M text prompts, and created 5M generated images. With user preference annotations from our dataset, we were able to train better preference alignment models. In addition, leveraging individual user preference, we investigated the performance of retrieval models and a vision-language model on personalized image retrieval and generative model recommendation. Finally, we propose an end-to-end framework for editing customized diffusion models in a latent weight space to align with individual user preferences. Our results demonstrate that the ImageGem dataset enables, for the first time, a new paradigm for generative model personalization.

📄 PDF Abstract BibTeX arXiv:2510.18433

Code (0)

등록된 구현이 없습니다.

Tasks

Image Retrieval

Similar Papers 제목 키워드 기반

3D-Aware Facial Landmark Detection via Multi-View Consistent Training on Synthetic Data

2023-01-01 · CVPR 2023 1 · Libing Zeng, Lele Chen, Wentao Bao, Zhong Li 외

Accurate facial landmark detection on wild images plays an essential role in human-computer interaction, entertainment, and medical applications. Existing approaches have limitations in enforcing 3D consistency while…

Facial Landmark DetectionImage GenerationNeural Rendering

WildFake: A Large-scale Challenging Dataset for AI-Generated Images Detection

2024-02-19 · Yan Hong, Jianfu Zhang

The extraordinary ability of generative models enabled the generation of images with such high quality that human beings cannot distinguish Artificial Intelligence (AI) generated images from real-life photographs. The de…

Monocular Human-Object Reconstruction in the Wild

2024-07-30 · Chaofan Huo, Ye Shi, Jingya Wang

Learning the prior knowledge of the 3D human-object spatial relation is crucial for reconstructing human-object interaction from images and understanding how humans interact with objects in 3D space. Previous works learn…

DiversityHuman-Object Interaction DetectionObjectObject Reconstruction

LiWi: Layering in the Wild

2026-05-14 · Yu He, Fang Li, Haoyang Tong, Lichen Ma 외 arxiv

Recent advances in generative models have empowered impressive layered image generation, yet their success is largely confined to graphic design domains. The layering of in-the-wild images remains an underexplored proble…

Image Generation

PhysX-Anything: Simulation-Ready Physical 3D Assets from Single Image

2025-11-17 · Ziang Cao, Fangzhou Hong, Zhaoxi Chen, Liang Pan 외 arxiv

3D modeling is shifting from static visual representations toward physical, articulated assets that can be directly used in simulation and interaction. However, most existing 3D generation methods overlook key physical a…

3D Generation