Enabling First-Order Gradient-Based Learning for Equilibrium Computation in Markets
Nils Kohring, Fabian Raoul Pieroth, Martin Bichler
摘要
Understanding and analyzing markets is crucial, yet analytical equilibrium solutions remain largely infeasible. Recent breakthroughs in equilibrium computation rely on zeroth-order policy gradient estimation. These approaches commonly suffer from high variance and are computationally expensive. The use of fully differentiable simulators would enable more efficient gradient estimation. However, the discrete allocation of goods in economic simulations is a non-differentiable operation. This renders the first-order Monte Carlo gradient estimator inapplicable and the learning feedback systematically misleading. We propose a novel smoothing technique that creates a surrogate market game, in which first-order methods can be applied. We provide theoretical bounds on the resulting bias which justifies solving the smoothed game instead. These bounds also allow choosing the smoothing strength a priori such that the resulting estimate has low variance. Furthermore, we validate our approach via numerous empirical experiments. Our method theoretically and empirically outperforms zeroth-order methods in approximation quality and computational efficiency.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Automated Design of Affine Maximizer Mechanisms in Dynamic SettingsMichael J. Curry, Vinzenz Thoma, Darshan Chakrabarti, Stephen McAleer 等AAAI 2024 · 被引用 13 次
- Auctionformer: A Unified Deep Learning Algorithm for Solving Equilibrium Strategies in Auction GamesKexin Huang, Ziqian Chen, Xue Wang, Chongming Gao 等ICML 2024 · 被引用 3 次
- Computing Perfect Bayesian Equilibria in Sequential Auctions with VerificationVinzenz Thoma, Vitor Bosshard, Sven SeukenAAAI 2025 · 被引用 3 次
- Scalable Neural Incentive Design with Parameterized Mean-Field ApproximationNathan Corecco, Batuhan Yardim, Vinzenz Thoma, Zebang Shen 等NeurIPS 2025 · 被引用 1 次
- Learning Bayesian Nash Equilibrium in Auction Games via Approximate Best ResponseKexin Huang, Ziqian Chen, Xue Wang, Chongming Gao 等ICML 2025
它引用的顶会 Paper5
- PlasticineLab: A Soft-Body Manipulation Benchmark with Differentiable PhysicsZhiao Huang, Yuanming Hu, Tao Du, Siyuan Zhou 等ICLR 2021 · 被引用 164 次
- Do Differentiable Simulators Give Better Policy Gradients?Hyung Ju Terry Suh, Max Simchowitz, Kaiqing Zhang, Russ TedrakeICML 2022 · 被引用 129 次
- On the Impossibility of Global Convergence in Multi-Loss OptimizationAlistair LetcherICLR 2021 · 被引用 33 次
- Systematically differentiating parametric discontinuitiesSai Praveen Bangaru, Jesse Michel, Kevin Mu, Gilbert Bernstein 等SIGGRAPH 2021 · 被引用 30 次
- Evolution Strategies for Approximate Solution of Bayesian GamesZun Li, Michael P. WellmanAAAI 2021 · 被引用 19 次
相关 Paper
- Adaptive-Gradient Policy Optimization: Enhancing Policy Learning in Non-Smooth Differentiable SimulationsFeng Gao, Liangzhi Shi, Shenao Zhang, Zhaoran Wang 等ICML 2024 · 被引用 7 次
- Approximating Nash Equilibria in Normal-Form Games via Stochastic OptimizationIan Gemp, Luke Marris, Georgios PiliourasICLR 2024 · 被引用 14 次
- Generalizing Stochastic Smoothing for Differentiation and Gradient EstimationFelix Petersen, Christian Borgelt, Aashwin Mishra, Stefano ErmonICML 2026 · 被引用 4 次
- Learning Individual Behavior in Agent-Based Models with Graph Diffusion NetworksFrancesco Cozzi, Marco Pangallo, Alan Perotti, André Panisson 等NeurIPS 2025 · 被引用 5 次
- On the Second-Order Convergence of Biased Policy Gradient AlgorithmsSiqiao Mu, Diego KlabjanICML 2024 · 被引用 4 次
