Lune

NeurIPS2025顶会

VarFlow: Proper Scoring-Rule Diffusion Distillation via Energy Matching

Huiyang Shao, Xin Xia, Yuxi Ren, Xing Wang, Xuefeng Xiao

2025年份
1被引次数

摘要

Diffusion models achieve remarkable generative performance but are hampered by slow, iterative inference. Model distillation seeks to train a fast student generator. Variational Score Distillation (VSD) offers a principled KL-divergence minimization framework for this task. This method cleverly avoids computing the teacher model's Jacobian, but its student gradient relies on the score of the student's own noisy marginal distribution, ∇ xt log p ϕ,t (x t ). VSD thus requires approximations, such as training an auxiliary network to estimate this score. These approximations can introduce biases, cause training instability, or lead to an incomplete match of the target distribution, potentially focusing on conditional means rather than broader distributional features. We introduce VarFlow, a novel distillation method based on a framework we term Score-Rule Variational Distillation (SRVD) framework. VarFlow trains a one-step generator g ϕ (z) by directly minimizing an energy distance (derived from the strictly proper energy score) between the student's induced noisy data distribution p ϕ,t (x t ) and the teacher's target noisy distribution q t (x t ). This objective is estimated entirely using samples from these two distributions. Crucially, VarFlow bypasses the need to compute or approximate the intractable student score. By directly matching the full noisy marginal distributions, VarFlow aims for a more comprehensive and robust alignment between student and teacher, offering an efficient and theoretically grounded path to high-fidelity one-step generation.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

它引用的顶会 Paper33

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖