Enabling First-Order Gradient-Based Learning for Equilibrium Computation in Markets
Nils Kohring, Fabian Raoul Pieroth, Martin Bichler
Abstract
Understanding and analyzing markets is crucial, yet analytical equilibrium solutions remain largely infeasible. Recent breakthroughs in equilibrium computation rely on zeroth-order policy gradient estimation. These approaches commonly suffer from high variance and are computationally expensive. The use of fully differentiable simulators would enable more efficient gradient estimation. However, the discrete allocation of goods in economic simulations is a non-differentiable operation. This renders the first-order Monte Carlo gradient estimator inapplicable and the learning feedback systematically misleading. We propose a novel smoothing technique that creates a surrogate market game, in which first-order methods can be applied. We provide theoretical bounds on the resulting bias which justifies solving the smoothed game instead. These bounds also allow choosing the smoothing strength a priori such that the resulting estimate has low variance. Furthermore, we validate our approach via numerous empirical experiments. Our method theoretically and empirically outperforms zeroth-order methods in approximation quality and computational efficiency.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 424cf5b2-7f05-41c3-9044-ec60af486869Cited by top-tier papers5
- Automated Design of Affine Maximizer Mechanisms in Dynamic SettingsMichael J. Curry, Vinzenz Thoma, Darshan Chakrabarti, Stephen McAleer et al.AAAI 2024 · 13 citations
- Auctionformer: A Unified Deep Learning Algorithm for Solving Equilibrium Strategies in Auction GamesKexin Huang, Ziqian Chen, Xue Wang, Chongming Gao et al.ICML 2024 · 3 citations
- Computing Perfect Bayesian Equilibria in Sequential Auctions with VerificationVinzenz Thoma, Vitor Bosshard, Sven SeukenAAAI 2025 · 3 citations
- Scalable Neural Incentive Design with Parameterized Mean-Field ApproximationNathan Corecco, Batuhan Yardim, Vinzenz Thoma, Zebang Shen et al.NeurIPS 2025 · 1 citation
- Learning Bayesian Nash Equilibrium in Auction Games via Approximate Best ResponseKexin Huang, Ziqian Chen, Xue Wang, Chongming Gao et al.ICML 2025
Builds on5
- PlasticineLab: A Soft-Body Manipulation Benchmark with Differentiable PhysicsZhiao Huang, Yuanming Hu, Tao Du, Siyuan Zhou et al.ICLR 2021 · 164 citations
- Do Differentiable Simulators Give Better Policy Gradients?Hyung Ju Terry Suh, Max Simchowitz, Kaiqing Zhang, Russ TedrakeICML 2022 · 129 citations
- On the Impossibility of Global Convergence in Multi-Loss OptimizationAlistair LetcherICLR 2021 · 33 citations
- Systematically differentiating parametric discontinuitiesSai Praveen Bangaru, Jesse Michel, Kevin Mu, Gilbert Bernstein et al.SIGGRAPH 2021 · 30 citations
- Evolution Strategies for Approximate Solution of Bayesian GamesZun Li, Michael P. WellmanAAAI 2021 · 19 citations
Related papers
- Adaptive-Gradient Policy Optimization: Enhancing Policy Learning in Non-Smooth Differentiable SimulationsFeng Gao, Liangzhi Shi, Shenao Zhang, Zhaoran Wang et al.ICML 2024 · 7 citations
- Approximating Nash Equilibria in Normal-Form Games via Stochastic OptimizationIan Gemp, Luke Marris, Georgios PiliourasICLR 2024 · 14 citations
- Generalizing Stochastic Smoothing for Differentiation and Gradient EstimationFelix Petersen, Christian Borgelt, Aashwin Mishra, Stefano ErmonICML 2026 · 4 citations
- Learning Individual Behavior in Agent-Based Models with Graph Diffusion NetworksFrancesco Cozzi, Marco Pangallo, Alan Perotti, André Panisson et al.NeurIPS 2025 · 5 citations
- On the Second-Order Convergence of Biased Policy Gradient AlgorithmsSiqiao Mu, Diego KlabjanICML 2024 · 4 citations
