From Predictors to Samplers via the Training Trajectory
Soumya Ram, Akhila Ram
摘要
Sampling from trained predictors is fundamental for interpretability and as a compute-light alternative to diffusion models, but local samplers struggle on the rugged, high-frequency functions such models learn. We observe that standard neural‑network training implicitly produces a coarse‑to‑fine sequence of models. Early checkpoints suppress high‑degree/ high‑frequency components (Boolean monomials; spherical harmonics under NTK), while later checkpoints restore detail. We exploit this by running a simple annealed sampler across the training trajectory, using early checkpoints for high‑mobility proposals and later ones for refinement. In the Boolean domain, this can turn the exponential bottleneck arising from rugged landscapes or needle gadgets into a near-linear one. In the continuous domain, under the NTK regime, this corresponds to smoothing under the NTK kernel. Requiring no additional compute, our method shows strong empirical gains across a variety of synthetic and real-world tasks, including constrained sampling tasks that diffusion models are unable to handle.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper22
- Practical and Asymptotically Exact Conditional Sampling in Diffusion ModelsLuhuan Wu, Brian L. Trippe, Christian A. Naesseth, David M. Blei 等NeurIPS 2023 · 被引用 276 次
- Hidden Progress in Deep Learning: SGD Learns Parities Near the Computational LimitBoaz Barak, Benjamin L. Edelman, Surbhi Goel, Sham M. Kakade 等NeurIPS 2022 · 被引用 220 次
- Infinite attention: NNGP and NTK for deep attention networksJiri Hron, Yasaman Bahri, Jascha Sohl-Dickstein, Roman NovakICML 2020 · 被引用 147 次
- Better by default: Strong pre-tuned MLPs and boosted trees on tabular dataDavid Holzmüller, Léo Grinsztajn, Ingo SteinwartNeurIPS 2024 · 被引用 141 次
- Design-Bench: Benchmarks for Data-Driven Offline Model-Based OptimizationBrandon Trabucco, Xinyang Geng, Aviral Kumar, Sergey LevineICML 2022 · 被引用 126 次
相关 Paper
- Progressive Inference-Time Annealing of Diffusion Models for Sampling from Boltzmann DensitiesTara Akhound-Sadegh, Jungyoon Lee, Joey Bose, Valentin De Bortoli 等NeurIPS 2025 · 被引用 28 次
- Diffusing Differentiable RepresentationsYash Savani, Marc Finzi, J. Zico KolterNeurIPS 2024 · 被引用 1 次
- Accelerated Parallel Tempering via Neural TransportsLeo Zhang, Peter Potaptchik, Jiajun He, Yuanqi Du 等ICLR 2026 · 被引用 14 次
- DISK: Differentiable Sparse Kernel Complex for Efficient Spatially-Variant ConvolutionZhizhen Wu, Zhe Cao, Yuchi HuoICLR 2026
- NTK-Guided Implicit Neural TeachingChen Zhang, Wei Zuo, Bingyang Cheng, Yikun Wang 等CVPR 2026 · 被引用 3 次
