FACM: Flow-Anchored Consistency Models
Yansong Peng, Kai Zhu, Yu Liu, Pingyu Wu, Hebei Li, Xiaoyan Sun, Feng Wu
摘要
Continuous-time Consistency Models (CMs) promise efficient few-step generation but face significant challenges with training instability. We argue this instability stems from a fundamental conflict: Training the network exclusively on a shortcut objective leads to the catastrophic forgetting of the instantaneous velocity field that defines the flow. Our solution is to explicitly anchor the model in the underlying flow, ensuring high trajectory fidelity during training. We introduce the Flow-Anchored Consistency Model (FACM), where a Flow Matching (FM) task serves as a dynamic anchor for the primary CM shortcut objective. Key to this Flow-Anchoring approach is a novel expanded time interval strategy that unifies optimization for a single model while decoupling the two tasks to ensure stable, architecturally-agnostic training. By distilling a pre-trained LightningDiT model, our method achieves a state-of-the-art FID of 1.32 with two steps (NFE=2) and 1.70 with just one step (NFE=1) on ImageNet 256256. To address the challenge of scalability, we develop a memory-efficient Chain-JVP that resolves key incompatibilities with FSDP. This method allows us to scale FACM training on a 14B parameter model (Wan 2.2), accelerating its Text-to-Image inference from 240 to 2-8 steps. Our code and pretrained models: https://github.com/ali-vilab/FACM.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Improved Mean Flows: On the Challenges of Fastforward Generative ModelsZhengyang Geng, Yiyang Lu, Zongze Wu, Eli Shechtman 等CVPR 2026 · 被引用 116 次
- Transition Models: Rethinking the Generative Learning ObjectiveZidong Wang, Yiyuan Zhang, Xiaoyu Yue, Xiangyu Yue 等CVPR 2026 · 被引用 32 次
- TwinFlow: Realizing One-step Generation on Large Models with Self-adversarial FlowsZhenglin Cheng, Peng Sun, Jianguo Li, Tao LinICLR 2026 · 被引用 17 次
- Towards Sequence Modeling Alignment between Tokenizer and Autoregressive ModelPingyu Wu, Kai Zhu, Yu Liu, Longxiang Tang 等ICLR 2026 · 被引用 16 次
- Flow Map Distillation Without DataShangyuan Tong, Nanye Ma, Saining Xie, Tommi S. JaakkolaCVPR 2026 · 被引用 13 次
它引用的顶会 Paper15
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- Scalable Diffusion Models with TransformersWilliam Peebles, Saining XieICCV 2023 · 被引用 5,568 次
- FlashAttention-2: Faster Attention with Better Parallelism and Work PartitioningTri DaoICLR 2024 · 被引用 2,600 次
- Mean Flows for One-step Generative ModelingZhengyang Geng, Mingyang Deng, Xingjian Bai, Zico Kolter 等NeurIPS 2025 · 被引用 628 次
相关 Paper
- CMT: Mid-Training for Efficient Learning of Consistency, Mean Flow, and Flow-Map ModelsZheyuan Hu, Chieh-Hsin Lai, Yuki Mitsufuji, Stefano ErmonICLR 2026 · 被引用 21 次
- Simplifying, Stabilizing and Scaling Continuous-time Consistency ModelsCheng Lu, Yang SongICLR 2025
- Large Scale Diffusion Distillation via Score-Regularized Continuous-Time ConsistencyKaiwen Zheng, Yuji Wang, Qianli Ma, Huayu Chen 等ICLR 2026 · 被引用 77 次
- Align Your Flow: Scaling Continuous-Time Flow Map DistillationAmirmojtaba Sabour, Sanja Fidler, Karsten KreisNeurIPS 2025 · 被引用 91 次
- AlphaFlow: Understanding and Improving MeanFlow ModelsHuijie Zhang, Aliaksandr Siarohin, Willi Menapace, Michael Vasilkovsky 等ICLR 2026 · 被引用 44 次
