Lune

ICLR2026Top-tier venue

Horizon Imagination: Efficient On-Policy Rollout in Diffusion World Models

Lior Cohen, Ofir Nabati, Kaixin Wang, Navdeep Kumar, Shie Mannor

2026Year

Abstract

We study diffusion-based world models for reinforcement learning, which offer high generative fidelity but face critical efficiency challenges in control. Current methods either require heavyweight models at inference or rely on highly sequential imagination, both of which impose prohibitive computational costs. We propose Horizon Imagination (HI), an on-policy imagination process for discrete stochastic policies that denoises multiple future observations in parallel. HI incorporates a stabilization mechanism and a novel sampling schedule that decouples the denoising budget from the effective horizon over which denoising is applied while also supporting fractional steps-per-frame budgets (sub-step budgets). Experiments on Atari 100K and Craftium show that our approach maintains control performance with a sub-step budget of half the denoising steps (i.e., 0.5 denoising steps per frame) and achieves superior generation quality under varied schedules. Code is available at https://github.com/leor-c/horizon-imagination.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext ddd10e75-58a5-45e4-9236-632fddcb51dd

Builds on14

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines