paper-with-me

Papers

Consistent-Inversion: Reverse Consistency Guidance for Structure-Preserving Visual Editing

2026-06-05 · Xiaocheng Lu, Jingcai Guo, Song Guo arxiv

Text-guided diffusion models have become effective tools for real-image visual editing, where the edited image must follow a target instruction while preserving editing-irrelevant structure. Most training-free editors rely on inversion: a source image is mapped to a noisy latent trajectory and the terminal latent is reused for target-prompt denoising. This reuse is useful for preservation, but it also couples source reconstruction and target editing. The resulting trajectory mismatch may either damage background/layout details or over-constrain the intended edit. This paper presents Consistent-Inversion, a training-free reverse consistency guidance framework for structure-preserving visual editing. Instead of treating the inverted source latent as a fixed initialization, Consistent-Inversion checks whether an intermediate target trajectory can be reversed toward the source inversion trajectory under the source prompt. To make this check well-defined, we construct an auxiliary target-side noise representation, perform source-guided reverse denoising, and use the resulting reverse consistency discrepancy as a correction signal for selected early target denoising steps. The method does not update model parameters, is compatible with inversion-based editors, and introduces only a small inference overhead when applied sparsely. Experiments on PIE-Bench show that Consistent-Inversion improves background and structural fidelity under a unified SD3.5 protocol while maintaining target-prompt alignment, and compatibility experiments further verify the same correction principle on classical Stable-Diffusion inversion pipelines.

📄 PDF Abstract BibTeX arXiv:2606.07145

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Robust Physics-Guided Diffusion for Full-Waveform Inversion

2026-03-17 · Jishen Peng, Enze Jiang, Zheng Ma, Xiongbin Yan arxiv

We develop a robust physics-guided diffusion framework for full-waveform inversion that combines a score-based generative prior with likelihood guidance computed through wave-equation simulations. We adopt a transport-ba…

Optimal Transport for Rectified Flow Image Editing: Unifying Inversion-Based and Direct Methods

2025-08-04 · Marian Lupascu, Mihai-Sorin Stupariu arxiv

Image editing in rectified flow models remains challenging due to the fundamental trade-off between reconstruction fidelity and editing flexibility. While inversion-based methods suffer from trajectory deviation, recent …

Image Editing

FlowAlign: Trajectory-Regularized, Inversion-Free Flow-based Image Editing

2025-05-29 · Jeongsol Kim, Yeobin Hong, Jong Chul Ye

Recent inversion-free, flow-based image editing methods such as FlowEdit leverages a pre-trained noise-to-image flow model such as Stable Diffusion 3, enabling text-driven manipulation by solving an ordinary differential…

Zero-shot Face Editing via ID-Attribute Decoupled Inversion

2025-10-13 · Yang Hou, Minggu Wang, Jianjun Zhao arxiv

Recent advancements in text-guided diffusion models have shown promise for general image editing via inversion techniques, but often struggle to maintain ID and structural consistency in real face editing tasks. To addre…

Image Editing

Abstract Sound Fusion with Unconditioned Inversion Model

2025-06-13 · Jing Liu, EnQi Lian

An abstract sound is defined as a sound that does not disclose identifiable real-world sound events to a listener. Sound fusion aims to synthesize an original sound and a reference sound to generate a novel sound that ex…

model