paper-with-me

홈 › Papers

SpotEdit: Evaluating Visually-Guided Image Editing Methods

2025-08-25 · Sara Ghazanfari, Wei-An Lin, Haitong Tian, Ersin Yumer arxiv

Visually-guided image editing, where edits are conditioned on both visual cues and textual prompts, has emerged as a powerful paradigm for fine-grained, controllable content generation. Although recent generative models have shown remarkable capabilities, existing evaluations remain simple and insufficiently representative of real-world editing challenges. We present SpotEdit, a comprehensive benchmark designed to systematically assess visually-guided image editing methods across diverse diffusion, autoregressive, and hybrid generative models, uncovering substantial performance disparities. To address a critical yet underexplored challenge, our benchmark includes a dedicated component on hallucination, highlighting how leading models, such as GPT-4o, often hallucinate the existence of a visual cue and erroneously perform the editing task. Our code and benchmark are publicly released at https://github.com/SaraGhazanfari/SpotEdit.

📄 PDF Abstract BibTeX arXiv:2508.18159

Code (0)

등록된 구현이 없습니다.

Tasks

Image Editing

Similar Papers 제목 키워드 기반

SpotEdit: Selective Region Editing in Diffusion Transformers

2025-12-26 · Zhibin Qin, Zhenxiong Tan, Zeqing Wang, Songhua Liu 외 arxiv

Diffusion Transformer models have significantly advanced image editing by encoding conditional images and integrating them into transformer layers. However, most edits involve modifying only small regions, while current …

Image Editing

GIE-Bench: Towards Grounded Evaluation for Text-Guided Image Editing

2025-05-16 · Yusu Qian, Jiasen Lu, Tsu-Jui Fu, Xinze Wang 외

Editing images using natural language instructions has become a natural and expressive way to modify visual content; yet, evaluating the performance of such models remains challenging. Existing evaluation approaches ofte…

Instruction FollowingMultiple-choicetext-guided-image-editingtext similarity

GuidedStyle: Attribute Knowledge Guided Style Manipulation for Semantic Face Editing

2020-12-22 · Xianxu Hou, Xiaokang Zhang, Linlin Shen, Zhihui Lai 외

Although significant progress has been made in synthesizing high-quality and visually realistic face images by unconditional Generative Adversarial Networks (GANs), there still lacks of control over the generation proces…

AttributeImage Generation

GaussEdit: Adaptive 3D Scene Editing with Text and Image Prompts

2025-09-30 · Zhenyu Shu, Junlong Yu, Kai Chao, Shiqing Xin 외 arxiv

This paper presents GaussEdit, a framework for adaptive 3D scene editing guided by text and image prompts. GaussEdit leverages 3D Gaussian Splatting as its backbone for scene representation, enabling convenient Region of…

3D scene Editing

Evaluating Image Editing with LLMs: A Comprehensive Benchmark and Intermediate-Layer Probing Approach

2026-03-20 · Shiqi Gao, Zitong Xu, Kang Fu, Huiyu Duan 외 arxiv

Evaluating text-guided image editing (TIE) methods remains a challenging problem, as reliable assessment should simultaneously consider perceptual quality, alignment with textual instructions, and preservation of origina…

Image Editing