paper-with-me

홈 › Papers

Semantic Granularity Navigation in Image Editing

2026-05-20 · Liangsi Lu, Minzhe Guo, Xuhang Chen, Yang Shi arxiv

Despite the generative capabilities of diffusion and flow models, real-image editing remains constrained by a persistent trade-off between semantic editability and structural fidelity. We trace a primary cause of this limitation to the implicit coupling of edit progress with model scale in existing paradigms. Under this coupling, stronger edits typically require visiting noisier states, which spends computation on destabilizing layout before the semantic change is well localized. We introduce NaviEdit, a training-free inference-time controller that decouples edit progress from model scale traversal through a strict self-consistency contract. NaviEdit operates at the rollout level and leaves the underlying pretrained model unchanged. It treats scale as a control input and reallocates a fixed step budget toward semantically responsive intermediate scales instead of destructive high-noise regimes. Experiments show positive average gains across compatible editors and flow backbones, supporting decoupling as a portable inference-time control principle.

📄 PDF Abstract BibTeX arXiv:2605.21190

Code (0)

등록된 구현이 없습니다.

Tasks

Image Editing

Similar Papers 제목 키워드 기반

Prioritized Semantic Learning for Zero-shot Instance Navigation

2024-03-18 · Xinyu Sun, Lizhao Liu, Hongyan Zhi, Ronghe Qiu 외

We study zero-shot instance navigation, in which the agent navigates to a specific object without using object annotations for training. Previous object navigation approaches apply the image-goal navigation (ImageNav) ta…

Language ModellingObject

Human-Aligned MLLM Judges for Fine-Grained Image Editing Evaluation: A Benchmark, Framework, and Analysis

2026-02-13 · Runzhou Liu, Hailey Weingord, Sejal Mittal, Prakhar Dungarwal 외 arxiv

Evaluating image editing models remains challenging due to the coarse granularity and limited interpretability of traditional metrics, which often fail to capture aspects important to human perception and intent. Such me…

Image Editing

Text-Vision Co-Instructed Image Editing

2026-06-15 · Chenxi Xie, Yuhui Wu, Qiaosi Yi, Lei Zhang arxiv

Existing image editing methods can be generally categorized into textual instruction-based and visual prompt-based ones. Textual instructions are semantically expressive, but are limited by the coarse granularity of spat…

Image ManipulationImage Editing

Meta-CoT: Enhancing Granularity and Generalization in Image Editing

2026-04-27 · Shiyi Zhang, Yiji Cheng, Tiankai Hang, Zijin Yin 외 arxiv

Unified multi-modal understanding/generative models have shown improved image editing performance by incorporating fine-grained understanding into their Chain-of-Thought (CoT) process. However, a critical question remain…

Image Editing

DAFNet: Dynamic Auxiliary Fusion for Sequential Model Editing in Large Language Models

2024-05-31 · Taolin Zhang, Qizhou Chen, Dongyang Li, Chengyu Wang 외

Recently, while large language models (LLMs) have demonstrated impressive results, they still suffer from hallucination, i.e., the generation of false information. Model editing is the task of fixing factual mistakes in …

HallucinationModel Editing