Reinforcement Learning of Sequential Price Mechanisms
Gianluca Brero, Alon Eden, Matthias Gerstgrasser, David C. Parkes, Duncan Rheingans-Yoo
摘要
We introduce the use of reinforcement learning for indirect mechanisms, working with the existing class of sequential price mechanisms, which generalizes both serial dictatorship and posted price mechanisms and essentially characterizes all strongly obviously strategyproof mechanisms. Learning an optimal mechanism within this class forms a partially-observable Markov decision process. We provide rigorous conditions for when this class of mechanisms is more powerful than simpler static mechanisms, for sufficiency or insufficiency of observation statistics for learning, and for the necessity of complex (deep) policies. We show that our approach can learn optimal or near-optimal mechanisms in several experimental settings.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- A Context-Integrated Transformer-Based Neural Network for Auction DesignZhijian Duan, Jingwu Tang, Yutong Yin, Zhe Feng 等ICML 2022 · 被引用 46 次
- Oracles & Followers: Stackelberg Equilibria in Deep Multi-Agent Reinforcement LearningMatthias Gerstgrasser, David C. ParkesICML 2023 · 被引用 27 次
- Learning to Mitigate AI Collusion on Economic PlatformsGianluca Brero, Eric Mibuari, Nicolas Lepore, David C. ParkesNeurIPS 2022 · 被引用 22 次
- Platform Behavior under Market Shocks: A Simulation Framework and Reinforcement-Learning Based StudyXintong Wang, Gary Qiurui Ma, Alon Eden, Clara Li 等WWW 2023 · 被引用 15 次
- Understanding Strategic Platform Entry and Seller Exploration: A Stackelberg ModelGarrett Seo, Xintong Wang, David C. ParkesWWW 2026
它引用的顶会 Paper1
相关 Paper
- Pessimism meets VCG: Learning Dynamic Mechanism Design via Offline Reinforcement LearningBoxiang Lyu, Zhaoran Wang, Mladen Kolar, Zhuoran YangICML 2022 · 被引用 9 次
- Computing Optimal Equilibria and Mechanisms via Learning in Zero-Sum Extensive-Form GamesBrian Hu Zhang, Gabriele Farina, Ioannis Anagnostides, Federico Cacciamani 等NeurIPS 2023 · 被引用 17 次
- Synthesizing Programmatic Policies that Inductively GeneralizeJeevana Priya Inala, Osbert Bastani, Zenna Tavares, Armando Solar-LezamaICLR 2020 · 被引用 54 次
- Mechanisms for a No-Regret Agent: Beyond the Common PriorModibo K. Camara, Jason D. Hartline, Aleck C. JohnsenFOCS 2020 · 被引用 5 次
- Deep Reinforcement Learning Finds Bayes-Nash Equilibrium in Competitive Newsvendor ProblemsKassian Köck, Fabian Raoul Pieroth, Martin BichlerICML 2026
