paper-with-me

Papers

Flow-Anchored Consistency Models

2025-07-04 · Yansong Peng, Kai Zhu, Yu Liu, Pingyu Wu, Hebei Li, Xiaoyan Sun, Feng Wu

Continuous-time Consistency Models (CMs) promise efficient few-step generation but face significant challenges with training instability. We argue this instability stems from a fundamental conflict: by training a network to learn only a shortcut across a probability flow, the model loses its grasp on the instantaneous velocity field that defines the flow. Our solution is to explicitly anchor the model in the underlying flow during training. We introduce the Flow-Anchored Consistency Model (FACM), a simple but effective training strategy that uses a Flow Matching (FM) task as an anchor for the primary CM shortcut objective. This Flow-Anchoring approach requires no architectural modifications and is broadly compatible with standard model architectures. By distilling a pre-trained LightningDiT model, our method achieves a state-of-the-art FID of 1.32 with two steps (NFE=2) and 1.76 with just one step (NFE=1) on ImageNet 256x256, significantly outperforming previous methods. This provides a general and effective recipe for building high-performance, few-step generative models. Our code and pretrained models: https://github.com/ali-vilab/FACM.

📄 PDF Abstract BibTeX arXiv:2507.03738

Code (1)

ali-vilab/FACM 공식 구현 pytorch

Tasks

Image Generation

Methods 이 논문이 사용한 방법론

Consistency Models 설명 없음

Similar Papers 제목 키워드 기반

Observability and Consistency Analysis for Visual-Inertial Navigation with Anchored Feature Parameterizations

2026-06-17 · Mitchell Cohen, Vassili Korotkine, James Richard Forbes arxiv

This paper presents an analysis of the observability and consistency properties of filtering-based visual-inertial navigation systems (VINS) that utilize anchored feature representations. The unobservable subspace of VIN…

FocalPolicy: Frequency-Optimized Chunking and Locally Anchored Flow Matching for Coherent Visuomotor Policy

2026-05-15 · Qian He, Zhenshuo Yang, Wenqi Liang, Chunhui Hao 외 arxiv

Visuomotor policies aim to learn complex manipulation tasks from expert demonstrations. However, generating smooth and coherent trajectories remains challenging, as it requires balancing proximal precision with distal fo…

Flow-Map GRPO: Reinforcement Learning for Few-Step Flow-Map Generators via Anchored Stochastic Composition

2026-07-01 · Zhiqi Li, Wen Zhang, Bo Zhu arxiv

Few-step flow-map generators, such as consistency models and MeanFlow, accelerate sampling by directly learning long-range transport maps between noise and data. However, these models are typically deterministic, which m…

Reinforcement Learning

STAGE: Storyboard-Anchored Generation for Cinematic Multi-shot Narrative

2025-12-13 · Peixuan Zhang, Zijian Jia, Kaiqi Liu, Shuchen Weng 외 arxiv

While recent advancements in generative models have achieved remarkable visual fidelity in video synthesis, creating coherent multi-shot narratives remains a significant challenge. To address this, keyframe-based approac…

Video Generation

AnchorEdit: Maintaining Temporal Consistency in Multi-turn Image Editing via Causal Memory

2026-06-10 · Hang Xu, Xiaoxiao Ma, Guohui Zhang, Yu Hu 외 arxiv

Multi-turn image editing is essential for iterative design, yet current models often struggle with identity drift and error accumulation over successive steps. While existing research leverages video priors for consisten…

Instruction FollowingCausal InferenceImage Editing