paper-with-me

홈 › Papers

Draw Your Mind: Personalized Generation via Condition-Level Modeling in Text-to-Image Diffusion Models

2025-08-05 · Hyungjin Kim, Seokho Ahn, Young-Duk Seo arxiv

Personalized generation in T2I diffusion models aims to naturally incorporate individual user preferences into the generation process with minimal user intervention. However, existing studies primarily rely on prompt-level modeling with large-scale models, often leading to inaccurate personalization due to the limited input token capacity of T2I diffusion models. To address these limitations, we propose DrUM, a novel method that integrates user profiling with a transformer-based adapter to enable personalized generation through condition-level modeling in the latent space. DrUM demonstrates strong performance on large-scale datasets and seamlessly integrates with open-source text encoders, making it compatible with widely used foundation T2I models without requiring additional fine-tuning.

📄 PDF Abstract BibTeX arXiv:2508.03481

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Adversarial speech for voice privacy protection from Personalized Speech generation

2024-01-22 · Shihao Chen, Liping Chen, Jie Zhang, KongAik Lee 외

The rapid progress in personalized speech generation technology, including personalized text-to-speech (TTS) and voice conversion (VC), poses a challenge in distinguishing between generated and real speech for human list…

Speaker Verificationtext-to-speechText to SpeechVoice Conversion

YNTP-100: A Benchmark for Your Next Token Prediction with 100 People

2025-10-16 · Shiyao Ding, Takayuki Ito arxiv

Large language models (LLMs) trained for general \textit{next-token prediction} often fail to generate responses that reflect how specific individuals communicate. Progress on personalized alignment is further limited by…

Response Generation

MindMelody: A Closed-Loop EEG-Driven System for Personalized Music Intervention

2026-05-02 · Yimeng Zhang, Yueru Sun, Haoyu Gu, Zhanpeng Jin arxiv

Driven by the escalating global burden of mental health conditions, music-based interventions have attracted significant attention as a non-invasive, cost-effective modality for emotion regulation and psychological stres…

Music Generation

TailorMind: Towards Preference-Aligned Multimodal Content Generation

2026-06-22 · Hengji Zhou, Ye Liu, Yufeng Liu, Si Wu 외 arxiv

Personalized content systems depend on available UGC and struggle when suitable content is absent, delayed, or costly to create. Although multimodal generators can synthesize content on demand, how to translate behaviora…

Collaborative Filteringmultimodal generation

From "What to Eat?" to Perfect Recipe: ChefMind's Chain-of-Exploration for Ambiguous User Intent in Recipe Recommendation

2025-09-22 · Yu Fu, Linyue Cai, Ruoyu Wu, Yong Zhao arxiv

Personalized recipe recommendation faces challenges in handling fuzzy user intent, ensuring semantic accuracy, and providing sufficient detail coverage. We propose ChefMind, a hybrid architecture combining Chain of Explo…