paper-with-me

홈 › Papers

ReFACT: Updating Text-to-Image Models by Editing the Text Encoder

2023-06-01 · Dana Arad, Hadas Orgad, Yonatan Belinkov

Our world is marked by unprecedented technological, global, and socio-political transformations, posing a significant challenge to text-to-image generative models. These models encode factual associations within their parameters that can quickly become outdated, diminishing their utility for end-users. To that end, we introduce ReFACT, a novel approach for editing factual associations in text-to-image models without relaying on explicit input from end-users or costly re-training. ReFACT updates the weights of a specific layer in the text encoder, modifying only a tiny portion of the model's parameters and leaving the rest of the model unaffected. We empirically evaluate ReFACT on an existing benchmark, alongside a newly curated dataset. Compared to other methods, ReFACT achieves superior performance in both generalization to related concepts and preservation of unrelated concepts. Furthermore, ReFACT maintains image generation quality, making it a practical tool for updating and correcting factual information in text-to-image models.

📄 PDF Abstract BibTeX arXiv:2306.00738

Code (1)

technion-cs-nlp/refact 공식 구현 pytorch

Tasks

Image Generation

Similar Papers 제목 키워드 기반

Source Prompt Disentangled Inversion for Boosting Image Editability with Diffusion Models

2024-03-17 · Ruibin Li, Ruihuang Li, Song Guo, Lei Zhang

Text-driven diffusion models have significantly advanced the image editing performance by using text prompts as inputs. One crucial step in text-driven image editing is to invert the original image into a latent noise co…

Image Generation

GaussCtrl: Multi-View Consistent Text-Driven 3D Gaussian Splatting Editing

2024-03-13 · Jing Wu, Jia-Wang Bian, Xinghui Li, Guangrun Wang 외

We propose GaussCtrl, a text-driven method to edit a 3D scene reconstructed by the 3D Gaussian Splatting (3DGS). Our method first renders a collection of images by using the 3DGS and edits them by using a pre-trained 2D …

3DGS

LocInv: Localization-aware Inversion for Text-Guided Image Editing

2024-05-02 · Chuanming Tang, Kai Wang, Fei Yang, Joost Van de Weijer

Large-scale Text-to-Image (T2I) diffusion models demonstrate significant generation capabilities based on textual prompts. Based on the T2I diffusion models, text-guided image editing research aims to empower users to ma…

Denoisingtext-guided-image-editing

Fast Multi-view Consistent 3D Editing with Video Priors

2025-11-28 · Liyi Chen, Ruihuang Li, Guowen Zhang, Pengfei Wang 외 arxiv

Text-driven 3D editing enables user-friendly 3D object or scene editing with text instructions. Due to the lack of multi-view consistency priors, existing methods typically resort to employing 2D generation or editing mo…

Video Generation

Dynamic Prompt Learning: Addressing Cross-Attention Leakage for Text-Based Image Editing

2023-09-27 · NeurIPS 2023 11 · Kai Wang, Fei Yang, Shiqi Yang, Muhammad Atif Butt 외

Large-scale text-to-image generative models have been a ground-breaking development in generative AI, with diffusion models showing their astounding ability to synthesize convincing images following an input text prompt.…

Prompt LearningText-based Image Editing