Markup-to-Image Diffusion Models with Scheduled Sampling
Yuntian Deng, Noriyuki Kojima, Alexander M. Rush
Abstract
Building on recent advances in image generation, we present a fully data-driven approach to rendering markup into images. The approach is based on diffusion models, which parameterize the distribution of data using a sequence of denoising operations on top of a Gaussian noise distribution. We view the diffusion denoising process as a sequential decision making process, and show that it exhibits compounding errors similar to exposure bias issues in imitation learning problems. To mitigate these issues, we adapt the scheduled sampling algorithm to diffusion training. We conduct experiments on four markup datasets: mathematical formulas (LaTeX), table layouts (HTML), sheet music (LilyPond), and molecular images (SMILES). These experiments each verify the effectiveness of the diffusion process and the use of scheduled sampling to fix generation issues. These results also show that the markup-to-image task presents a useful controlled compositional setting for diagnosing and analyzing generative image models.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 69d29db6-8b57-4839-9e4d-a78317d46f5dCited by top-tier papers4
- Multi-Step Denoising Scheduled Sampling: Towards Alleviating Exposure Bias for Diffusion ModelsZhiyao Ren, Yibing Zhan, Liang Ding, Gaoang Wang et al.AAAI 2024 · 15 citations
- NeuralOS: Towards Simulating Operating Systems via Neural Generative ModelsLuke Rivard, Sun Sun, Hongyu Guo, Wenhu Chen et al.ICLR 2026 · 13 citations
- Contrast-augmented Diffusion Model with Fine-grained Sequence Alignment for Markup-to-Image GenerationGuojin Zhong, Jin Yuan, Pan Wang, Kailun Yang et al.ACM MM 2023 · 7 citations
- Bidirectional Noise Injection: Enhancing Diffusion Models via Coordinated Input-Output PerturbationTianyi Zheng, Jiayang Gao, Peng-Tao Jiang, Fengxiang Yang et al.AAAI 2026
Builds on10
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
- Zero-Shot Text-to-Image GenerationAditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray et al.ICML 2021 · 6,356 citations
Related papers
- Why Masking Diffusion Works: Condition on the Jump Schedule for Improved Discrete DiffusionAlan Nawzad Amin, Nate Gruver, Andrew Gordon WilsonNeurIPS 2025 · 18 citations
- Is Your Diffusion Model Actually Denoising?Daniel Pfrommer, Zehao Dou, Christopher Scarvelis, Max Simchowitz et al.NeurIPS 2025 · 1 citation
- Align Your Steps: Optimizing Sampling Schedules in Diffusion ModelsAmirmojtaba Sabour, Sanja Fidler, Karsten KreisICML 2024 · 74 citations
- Inverse Problem Sampling in Latent Space Using Sequential Monte CarloIdan Achituve, Hai Victor Habi, Amir Rosenfeld, Arnon Netzer et al.ICML 2025
- Anti-Exposure Bias in Diffusion ModelsJunyu Zhang, Daochang Liu, Eunbyung Park, Shichao Zhang et al.ICLR 2025
