Diffusion Tree Sampling: Scalable inference‑time alignment of diffusion models
Vineet Jain, Kusha Sareen, Mohammad Pedramfar, Siamak Ravanbakhsh
摘要
Adapting a pretrained diffusion model to new objectives at inference time remains an open problem in generative modeling. Existing steering methods suffer from inaccurate value estimation, especially at high noise levels, which biases guidance. Moreover, information from past runs is not reused to improve sample quality, resulting in inefficient use of compute. Inspired by the success of Monte Carlo Tree Search, we address these limitations by casting inference-time alignment as a search problem that reuses past computations. We introduce a tree-based approach that samples from the reward-aligned target density by propagating terminal rewards back through the diffusion chain and iteratively refining value estimates with each additional generation. Our proposed method, Diffusion Tree Sampling (DTS), produces asymptotically exact samples from the target distribution in the limit of infinite rollouts, and its greedy variant, Diffusion Tree Search (DTS), performs a global search for high reward samples. On MNIST and CIFAR-10 class-conditional generation, DTS matches the FID of the best-performing baseline with up to less compute. In text-to-image generation and language completion tasks, DTS effectively searches for high reward samples that match best-of-N with up to less compute. By reusing information from previous generations, we get an anytime algorithm that turns additional compute into steadily better samples, providing a scalable approach for inference-time alignment of diffusion models.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- Fast Solvers for Discrete Diffusion Models: Theory and Applications of High-Order AlgorithmsYinuo Ren, Haoxuan Chen, Yuchen Zhu, Wei Guo 等NeurIPS 2025 · 被引用 51 次
- Inference-time Physics Alignment of Video Generative Models with Latent World ModelsJianhao Yuan, Xiaofeng Zhang, Felix Friedrich, Nicolas Beltran-Velez 等CVPR 2026 · 被引用 32 次
- Meta Flow Maps enable scalable reward alignmentPeter Potaptchik, Adhi Saravanan, Abbas Mammadov, Alvaro Prat 等ICML 2026 · 被引用 25 次
- Inference-Time Scaling of Discrete Diffusion Models via Importance Weighting and Optimal Proposal DesignZijing Ou, Chinmay Pani, Yingzhen LiICLR 2026 · 被引用 14 次
- From Scale to Speed: Adaptive Test-Time Scaling for Image EditingXiangyan Qu, Zhenlong Yuan, Jing Tang, Rui Chen 等CVPR 2026 · 被引用 8 次
它引用的顶会 Paper41
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
相关 Paper
- GLASS Flows: Efficient Inference for Reward Alignment of Flow and Diffusion ModelsPeter Holderrieth, Uriel Singer, Tommi Jaakkola, Ricky T. Q. Chen 等ICLR 2026 · 被引用 7 次
- Test-time Alignment of Diffusion Models without Reward Over-optimizationSunwoo Kim, Minkyu Kim, Dongmin ParkICLR 2025
- Controllable Graph Generation with Diffusion Models via Inference-Time Tree Search GuidanceJiachi Zhao, Zehong Wang, Yamei Liao, Chuxu Zhang 等WWW 2026 · 被引用 4 次
- Inference-time scaling of diffusion models through classical searchXiangcheng Zhang, Haowei Lin, Haotian Ye, James Y. Zou 等ICLR 2026 · 被引用 57 次
- Diamond Maps: Efficient Reward Alignment via Stochastic Flow MapsPeter Holderrieth, Douglas Chen, Luca Eyring, Ishin Shah 等ICML 2026 · 被引用 18 次
