paper-with-me

홈 › Papers

Geometric Disentanglement of Text Embeddings for Subject-Consistent Text-to-Image Generation using A Single Prompt

2025-12-18 · Shangxun Li, Youngjung Uh arxiv

Text-to-image diffusion models excel at generating high-quality images from natural language descriptions but often fail to preserve subject consistency across multiple outputs, limiting their use in visual storytelling. Existing approaches rely on model fine-tuning or image conditioning, which are computationally expensive and require per-subject optimization. 1Prompt1Story, a training-free approach, concatenates all scene descriptions into a single prompt and rescales token embeddings, but it suffers from semantic leakage, where embeddings across frames become entangled, causing text misalignment. In this paper, we propose a simple yet effective training-free approach that addresses semantic entanglement from a geometric perspective by refining text embeddings to suppress unwanted semantics. Extensive experiments prove that our approach significantly improves both subject consistency and text alignment over existing baselines.

📄 PDF Abstract BibTeX arXiv:2512.16443

Code (0)

등록된 구현이 없습니다.

Tasks

Text-to-Image GenerationVisual Storytelling

Similar Papers 제목 키워드 기반

Decoupled Textual Embeddings for Customized Image Generation

2023-12-19 · Yufei Cai, Yuxiang Wei, Zhilong Ji, Jinfeng Bai 외

Customized text-to-image generation, which aims to learn user-specified concepts with a few images, has drawn significant attention recently. However, existing methods usually suffer from overfitting issues and entangle …

AttributeDisentanglementImage GenerationText to Image Generation+2

Neutralizing Gender Bias in Word Embedding with Latent Disentanglement and Counterfactual Generation

2020-04-07 · Seungjae Shin, Kyungwoo Song, JoonHo Jang, Hyemi Kim 외

Recent research demonstrates that word embeddings, trained on the human-generated corpus, have strong gender biases in embedding spaces, and these biases can result in the discriminative results from the various downstre…

counterfactualDisentanglementSentiment AnalysisWord Embeddings

Understanding and Enforcing Weight Disentanglement in Task Arithmetic

2026-04-18 · Shangge Liu, Yuehan Yin, Lei Wang, Qi Fan 외 arxiv

Task arithmetic provides an efficient, training-free way to edit pre-trained models, yet lacks a fundamental theoretical explanation for its success. The existing concept of ``weight disentanglement" describes the ideal …

JE-IRT: A Geometric Lens on LLM Abilities through Joint Embedding Item Response Theory

2025-09-26 · Louie Hong Yao, Nicholas Jarvis, Tiffany Zhan, Saptarshi Ghosh 외 arxiv

Standard LLM evaluation practices compress diverse abilities into single scores, obscuring their inherently multidimensional nature. We present JE-IRT, a geometric item-response framework that embeds both LLMs and questi…

Global Facts

Neutralizing Gender Bias in Word Embeddings with Latent Disentanglement and Counterfactual Generation

2020-11-01 · Findings of the Association for Computational Linguistics 2020 · Seungjae Shin, Kyungwoo Song, JoonHo Jang, Hyemi Kim 외

Recent research demonstrates that word embeddings, trained on the human-generated corpus, have strong gender biases in embedding spaces, and these biases can result in the discriminative results from the various downstre…

counterfactualDisentanglementWord Embeddings