MarioGPT: Open-Ended Text2Level Generation through Large Language Models
Shyam Sudhakaran, Miguel González Duque, Matthias Freiberger, Claire Glanois, Elias Najarro, Sebastian Risi
Abstract
Procedural Content Generation (PCG) algorithms provide a technique to generate complex and diverse environments in an automated way. However, while generating content with PCG methods is often straightforward, generating meaningful content that reflects specific intentions and constraints remains challenging. Furthermore, many PCG algorithms lack the ability to generate content in an open-ended manner. Recently, Large Language Models (LLMs) have shown to be incredibly effective in many diverse domains. These trained LLMs can be fine-tuned, re-using information and accelerating training for new tasks. In this work, we introduce MarioGPT, a fine-tuned GPT2 model trained to generate tile-based game levels, in our case Super Mario Bros levels. We show that MarioGPT can not only generate diverse levels, but can be text-prompted for controllable level generation, addressing one of the key challenges of current PCG techniques. As far as we know, MarioGPT is the first text-to-level model. We also combine MarioGPT with novelty search, enabling it to generate diverse levels with varying play-style dynamics (i.e. player paths). This combination allows for the open-ended generation of an increasingly diverse range of content.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8fb1e556-ae7c-4558-8f6a-62ed53a51174Cited by top-tier papers11
- Darwin Gödel Machine: Open-Ended Evolution of Self-Improving AgentsJenny Zhang, Shengran Hu, Cong Lu, Robert Tjarko Lange et al.ICLR 2026 · 101 citations
- Quality-Diversity through AI FeedbackHerbie Bradley, Andrew Dai, Hannah Benita Teufel, Jenny Zhang et al.ICLR 2024 · 43 citations
- GAVEL: Generating Games via Evolution and Language ModelsGraham Todd, Alexander Padula, Matthew Stephenson, Éric Piette et al.NeurIPS 2024 · 34 citations
- DreamGarden: A Designer Assistant for Growing Games from a Single PromptSam Earle, Samyak Parajuli, Andrzej Banburski-FaheyCHI 2025 · 14 citations
- FactorSim: Generative Simulation via Factorized RepresentationFan-Yun Sun, S. I. Harini, Angela Yi, Yihan Zhou et al.NeurIPS 2024 · 11 citations
Builds on3
- Memorization Without Overfitting: Analyzing the Training Dynamics of Large Language ModelsKushal Tirumala, Aram H. Markosyan, Luke Zettlemoyer, Armen AghajanyanNeurIPS 2022 · 304 citations
- Quantifying Memorization Across Neural Language ModelsNicholas Carlini, Daphne Ippolito, Matthew Jagielski, Katherine Lee et al.ICLR 2023 · 158 citations
- Illuminating Mario Scenes in the Latent Space of a Generative Adversarial NetworkMatthew C. Fontaine, Ruilin Liu, Ahmed Khalifa, Jignesh Modi et al.AAAI 2021 · 98 citations
Related papers
- Controllable Procedural Generation of LandscapesJia-Hong Liu, Shao-Kui Zhang, Chuyue Zhang, Song-Hai ZhangACM MM 2024 · 6 citations
- Text2VRScene: Exploring the Framework of Automated Text-driven Generation System for VR ExperienceZhizhuo Yin, Yuyang Wang, Theodoros Papatheodorou, Pan HuiIEEE VR 2024 · 25 citations
- Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement LearningSimon Zhai, Hao Bai, Zipeng Lin, Jiayi Pan et al.NeurIPS 2024 · 214 citations
- TLPO: Token-Level Policy Optimization for Mitigating Language Confusion in Large Language ModelsJinho Choo, JunSeung Lee, Jimyeong Kim, Yeeho Song et al.ACL 2026
- GAPO: Learning Preferential Prompt through Generative Adversarial Policy OptimizationZhouhong Gu, Xingzhou Chen, Xiaoran Shi, Tao Wang et al.ACL 2025
