paper-with-me

홈 › Papers

Safe Few-Step Generation via Velocity Editing

2026-06-22 · Yujin Choi, Jaehong Yoon arxiv

Flow matching has recently emerged as a strong paradigm for state-of-the-art text-to-image (T2I) generation, enabling high-quality generation with a small number of sampling steps. As these models are increasingly integrated into real-world applications, ensuring safe and non-sensitive content generation has become a critical requirement. However, adapting safety and concept removal methods to this new generation framework remains an open challenge. Specifically, prior methods largely rely on iterative trajectory steering across a number of denoising steps or on CLIP-centric prompt embedding manipulation. These design assumptions pose fundamental bottlenecks for safety in flow matching-based T2I generation, where limited sampling steps constrain iterative correction and modern context-aware text encoders diminish the effectiveness of embedding-level interventions. In this paper, we propose VESFlow, a training-free safety method tailored to flow matching with extremely few sampling steps. Leveraging the fact that flow matching models learn the marginal velocity, we directly edit the velocity field via a safe-conditional posterior. VESFlow steers the trajectory toward safe outputs while leaving the conditioning prompt unchanged. Building on the observation that VESFlow leaves outputs unchanged under benign prompts, we further introduce a risk score-based filtering that bypasses velocity editing to reduce computational cost while preserving benign prompt generation. Based on this filtering, we propose VESFlow+, a stronger variant of VESFlow that not only edits the velocity toward the safe direction, but also pushes it away from the unsafe direction. Experimental results show that VESFlow+ removes the target concept, reducing the attack success rate by NudeNet to 6.3% on Ring-A-Bell and 6.8% on MMA-Diffusion on the 4-step MeanFlow model, while preserving fidelity on benign prompts.

📄 PDF Abstract BibTeX arXiv:2606.23267

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

VAGS: Velocity Adaptive Guidance Scale for Image Editing and Generation

2026-05-15 · Yan Luo, Ahmadou Aidara, Jingyi Lu, Jeremy Moebel 외 arxiv

Classifier-free guidance (CFG) is the primary control over how strongly text semantics move a flow-based sampler, yet standard practice holds its scale fixed across the entire ODE trajectory. This is a fundamental mismat…

Image Editing

BiFM: Bidirectional Flow Matching for Few-Step Image Editing and Generation

2026-03-26 · Yasong Dai, Zeeshan Hayder, David Ahmedt-Aristizabal, Hongdong Li arxiv

Recent diffusion and flow matching models have demonstrated strong capabilities in image generation and editing by progressively removing noise through iterative sampling. While this enables flexible inversion for semant…

Image GenerationImage Editing

FastFlow: Accelerating The Generative Flow Matching Models with Bandit Inference

2026-02-11 · Divya Jyoti Bajpai, Dhruv Bhardwaj, Soumya Roy, Tejas Duseja 외 arxiv

Flow-matching models deliver state-of-the-art fidelity in image and video generation, but the inherent sequential denoising process renders them slower. Existing acceleration methods like distillation, trajectory truncat…

Video GenerationImage Generation

Free Lunch for Stabilizing Rectified Flow Inversion

2026-02-12 · Chenru Wang, Beier Zhu, Chi Zhang arxiv

Rectified-Flow (RF)-based generative models have recently emerged as strong alternatives to traditional diffusion models, demonstrating state-of-the-art performance across various tasks. By learning a continuous velocity…

Image Reconstruction

Chain-of-Jailbreak Attack for Image Generation Models via Editing Step by Step

2024-10-04 · Wenxuan Wang, Kuiyi Gao, Youliang Yuan, Jen-tse Huang 외

Text-based image generation models, such as Stable Diffusion and DALL-E 3, hold significant potential in content creation and publishing workflows, making them the focus in recent years. Despite their remarkable capabili…

Image Generation