paper-with-me

Papers

Edit Fidelity Field: Semantics-Aware Region Isolation for Training-Free Scene Text Editing

2026-04-19 · Guandong Li, Mengxia Ye arxiv

Scene text editing (STE) has achieved remarkable progress in accurately rendering target text through diffusion-based methods. However, we identify a critical yet overlooked problem: edit spillover -- when editing a target text region, existing methods inadvertently modify non-target regions, particularly neighboring text. Through systematic evaluation on 50 real-world scenes across four categories, we reveal that state-of-the-art diffusion editing models exhibit a spillover rate of 94%, meaning nearly all non-target text regions are altered during editing. To address this, we propose the Edit Fidelity Field (EFF), a semantics-aware continuous field that controls per-pixel editing fidelity. Unlike binary masks, EFF leverages OCR-detected text regions to construct a four-zone field: Edit Core (fully editable), Transition Zone (smooth decay), Protected Zone (non-target text, explicitly locked), and Background (strictly preserved). EFF operates as a training-free, model-agnostic post-processing module applicable to any diffusion-based STE method. We further propose per-region spillover quantification, a novel evaluation protocol that measures edit leakage at each non-target text region individually. Experiments demonstrate that EFF reduces spillover rate from 94% to 25% while improving non-target region preservation by +91.4 dB PSNR.

📄 PDF Abstract BibTeX arXiv:2604.17500

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

FENeRF: Face Editing in Neural Radiance Fields

2021-11-30 · CVPR 2022 1 · Jingxiang Sun, Xuan Wang, Yong Zhang, Xiaoyu Li 외

Previous portrait image generation methods roughly fall into two categories: 2D GANs and 3D-aware GANs. 2D GANs can generate high fidelity portraits but with low view consistency. 3D-aware GAN methods can maintain view c…

3D-Aware Image SynthesisImage Generation

HierEdit: Region-Aware Hierarchical Diffusion for Efficient High-Resolution Editing

2026-05-17 · Yuyao Zhang, Alexander Huang-Menders, Yu-Wing Tai arxiv

High-resolution image editing is essential for professional and creative applications, yet existing multimodal diffusion-based editors remain computationally inefficient and constrained to relatively low resolutions. Cur…

Image Editing

SemFaceEdit: Semantic Face Editing on Generative Radiance Manifolds

2025-06-28 · Shashikant Verma, Shanmuganathan Raman

Despite multiple view consistency offered by 3D-aware GAN techniques, the resulting images often lack the capacity for localized editing. In response, generative radiance manifolds emerge as an efficient approach for con…

Disentanglement

FAME: Fairness-aware Attention-modulated Video Editing

2025-10-27 · Zhangkai Wu, Xuhui Fan, Zhongyuan Xie, Kaize Shi 외 arxiv

Training-free video editing (VE) models tend to fall back on gender stereotypes when rendering profession-related prompts. We propose \textbf{FAME} for \textit{Fairness-aware Attention-modulated Video Editing} that mitig…

Region-Aware Diffusion for Zero-shot Text-driven Image Editing

2023-02-23 · Nisha Huang, Fan Tang, WeiMing Dong, Tong-Yee Lee 외

Image manipulation under the guidance of textual descriptions has recently received a broad range of attention. In this study, we focus on the regional editing of images with the guidance of given text prompts. Different…

Image Manipulation