G-Sim: Generative Simulations with Large Language Models and Gradient-Free Calibration
Samuel Holt, Max Ruiz Luyten, Antonin Berthon, Mihaela van der Schaar
Abstract
Constructing robust simulators is essential for asking "what if?" questions and guiding policy in critical domains like healthcare and logistics. However, existing methods often struggle, either failing to generalize beyond historical data or, when using Large Language Models (LLMs), suffering from inaccuracies and poor empirical alignment. We introduce G-Sim, a hybrid framework that automates simulator construction by synergizing LLM-driven structural design with rigorous empirical calibration. G-Sim employs an LLM in an iterative loop to propose and refine a simulator's core components and causal relationships, guided by domain knowledge. This structure is then grounded in reality by estimating its parameters using flexible calibration techniques. Specifically, G-Sim can leverage methods that are both likelihood-free and gradient-free with respect to the simulator, such as gradient-free optimization for direct parameter estimation or simulation-based inference for obtaining a posterior distribution over parameters. This allows it to handle non-differentiable and stochastic simulators. By integrating domain priors with empirical evidence, G-Sim produces reliable, causallyinformed simulators, mitigating data-inefficiency and enabling robust system-level interventions for complex decision-making.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on15
- Retrieval-Augmented Generation for Knowledge-Intensive NLP TasksPatrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni et al.NeurIPS 2020 · 19,162 citations
- Evaluating the World Model Implicit in a Generative ModelKeyon Vafa, Justin Y. Chen, Ashesh Rambachan, Jon M. Kleinberg et al.NeurIPS 2024 · 166 citations
- Causal Transformer for Estimating Counterfactual OutcomesValentyn Melnychuk, Dennis Frauen, Stefan FeuerriegelICML 2022 · 146 citations
- WorldCoder, a Model-Based LLM Agent: Building World Models by Writing Code and Interacting with the EnvironmentHao Tang, Darren Key, Kevin EllisNeurIPS 2024 · 123 citations
- Reasoning with Language Model is Planning with World ModelShibo Hao, Yi Gu, Haodi Ma, Joshua Jiahua Hong et al.EMNLP 2023 · 109 citations
Related papers
- Diagnostic-Guided Dynamic Profile Optimization for LLM-based User Simulators in Sequential RecommendationHongyang Liu, Zhu Sun, Tianjun Wei, Yan Wang et al.AAAI 2026 · 4 citations
- Mirroring Users: Towards Building Preference-aligned User Simulator with User Feedback in RecommendationTianjun Wei, Huizhong Guo, Yingpeng Du, Zhu Sun et al.ACL 2026 · 4 citations
- Towards Synthetic Trace Generation of Modeling Operations using In-Context Learning ApproachVittoriano Muttillo, Claudio Di Sipio, Riccardo Rubei, Luca Berardinelli et al.ASE 2024 · 1 citation
- LILO: Bayesian Optimization with Natural Language FeedbackKatarzyna Kobalczyk, Zhiyuan Lin, Benjamin Letham, Zhuokai Zhao et al.ICML 2026 · 2 citations
- Valid Inference with Imperfect Synthetic DataYewon Byun, Shantanu Gupta, Zachary C. Lipton, Rachel Leah Childers et al.NeurIPS 2025 · 8 citations
