paper-with-me

홈 › Papers

h-Flow: Flexible Flow-based Image Editing via Doob's h-Transform

2026-07-12 · Zehui Guo, Zhen Wang, Junwei Shu, Yang Li, Changbo Wang, Long Chen arxiv

Editing images with pre-trained text-to-image flow models typically requires carefully balancing target alignment with the desired prompt and source consistency with the original image. Existing approaches either rely on inversion-based pipelines or heuristic source-to-target trajectory constructions, which often depend on architecture-specific designs or are sensitive to hyperparameters. In this paper, we propose h-Flow, a training-free and theoretically grounded flow-based editing framework. Inspired by Doob's $h$-Transform, we reformulate image editing as conditional generation under multiple terminal events corresponding to source consistency and target alignment. We first extend the classical $h$-Transform from SDE-based models to the deterministic RF framework by constructing an equivalent SDE with identical marginals. Within this formulation, we design dedicated $h$-functions for source consistency and target alignment, yielding closed-form reconstruction guidance and velocity-based semantic editing signals. We further introduce a velocity orthogonal decomposition to decouple reconstruction and editing directions, enabling a controllable trade-off between the two objectives. Extensive experiments demonstrate that h-Flow achieves effective, robust, and flexible editing across diverse scenarios. The code will be released soon.

📄 PDF Abstract BibTeX arXiv:2607.10800

Code (0)

등록된 구현이 없습니다.

Tasks

Image Editing

Similar Papers 제목 키워드 기반

h-Edit: Effective and Flexible Diffusion-Based Editing via Doob's h-Transform

2025-03-04 · CVPR 2025 1 · Toan Nguyen, Kien Do, Duc Kieu, Thin Nguyen

We introduce a theoretical framework for diffusion-based image editing by formulating it as a reverse-time bridge modeling problem. This approach modifies the backward process of a pretrained diffusion model to construct…

Unveil Inversion and Invariance in Flow Transformer for Versatile Image Editing

2024-11-24 · CVPR 2025 1 · Pengcheng Xu, Boyuan Jiang, Xiaobin Hu, Donghao Luo 외

Leveraging the large generative prior of the flow transformer for tuning-free image editing requires authentic inversion to project the image into the model's domain and a flexible invariance control mechanism to preserv…

Mage-Flow: An Efficient Native-Resolution Foundation Model for Image Generation and Editing

2026-07-21 · Xinjie Zhang, Peng Zhang, Shicheng Zheng, Jinghao Guo 외 arxiv

Large-scale visual generators are increasingly capable but costly to train, fine-tune, and deploy. We introduce Mage-Flow, a compact 4B-scale generative stack for efficient text-to-image generation and instruction-based …

Text-to-Image GenerationImage Editing

Towards Training-Free Scene Text Editing

2026-03-25 · Yubo Li, Xugong Qin, Peng Zhang, Hailun Lin 외 arxiv

Scene text editing seeks to modify textual content in natural images while maintaining visual realism and semantic consistency. Existing methods often require task-specific training or paired data, limiting their scalabi…

BiFM: Bidirectional Flow Matching for Few-Step Image Editing and Generation

2026-03-26 · Yasong Dai, Zeeshan Hayder, David Ahmedt-Aristizabal, Hongdong Li arxiv

Recent diffusion and flow matching models have demonstrated strong capabilities in image generation and editing by progressively removing noise through iterative sampling. While this enables flexible inversion for semant…

Image GenerationImage Editing