paper-with-me

홈 › Papers

LatentEdit: Adaptive Latent Control for Consistent Semantic Editing

2025-08-30 · Siyi Liu, Weiming Chen, Yushun Tang, Zhihai He arxiv

Diffusion-based Image Editing has achieved significant success in recent years. However, it remains challenging to achieve high-quality image editing while maintaining the background similarity without sacrificing speed or memory efficiency. In this work, we introduce LatentEdit, an adaptive latent fusion framework that dynamically combines the current latent code with a reference latent code inverted from the source image. By selectively preserving source features in high-similarity, semantically important regions while generating target content in other regions guided by the target prompt, LatentEdit enables fine-grained, controllable editing. Critically, the method requires no internal model modifications or complex attention mechanisms, offering a lightweight, plug-and-play solution compatible with both UNet-based and DiT-based architectures. Extensive experiments on the PIE-Bench dataset demonstrate that our proposed LatentEdit achieves an optimal balance between fidelity and editability, outperforming the state-of-the-art method even in 8-15 steps. Additionally, its inversion-free variant further halves the number of neural function evaluations and eliminates the need for storing any intermediate variables, substantially enhancing real-time deployment efficiency.

📄 PDF Abstract BibTeX arXiv:2509.00541

Code (0)

등록된 구현이 없습니다.

Tasks

Image Editing

Similar Papers 제목 키워드 기반

LatentEditor: Text Driven Local Editing of 3D Scenes

2023-12-14 · Umar Khalid, Hasan Iqbal, Nazmul Karim, Jing Hua 외

While neural fields have made significant strides in view synthesis and scene reconstruction, editing them poses a formidable challenge due to their implicit encoding of geometry and texture information from multi-view i…

3D scene EditingDenoisingNeRF

Controllable Affective Generation via Latent Vector Steering

2026-08-26 · Xixian Yong, Siyuan Chang, Yingying Zhang, Xian Wu 외 arxiv

Large Language Models (LLMs) often produce emotionally flattened responses after alignment, limiting their effectiveness in affect-sensitive applications. In this paper, we propose EmoVec, a lightweight framework for con…

Continuous Control

IA-FaceS: A Bidirectional Method for Semantic Face Editing

2022-03-24 · Wenjing Huang, Shikui Tu, Lei Xu

Semantic face editing has achieved substantial progress in recent years. Known as a growingly popular method, latent space manipulation performs face editing by changing the latent code of an input face to liberate users…

Attribute

IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

2024-10-17 · Xiaoyu Chen, Junliang Guo, Tianyu He, Chuheng Zhang 외

We introduce Image-GOal Representations (IGOR), aiming to learn a unified, semantically consistent action space across human and various robots. Through this unified latent action space, IGOR enables knowledge transfer a…

Transfer Learning

Controlla: Learning Controllability via Graph-Constrained Latent Geometry

2026-05-15 · Jamuna S. Murthy, Amin Karimi Monsefi, Rajiv Ramnath arxiv

Controllable multimodal generation is commonly formulated as an inference-time conditioning problem using prompts, guidance, or auxiliary modules. While effective, such approaches do not explicitly structure how semantic…

multimodal generation