Multiple Mean-Payoff Optimization Under Local Stability Constraints
David Klaska, Antonín Kucera, Vojtech Kur, Vít Musil, Vojtech Rehák
Abstract
The long-run average payoff per transition (mean payoff) is the main tool for specifying the performance and dependability properties of discrete systems. The problem of constructing a controller (strategy) simultaneously optimizing several mean payoffs has been deeply studied for stochastic and game-theoretic models. One common issue of the constructed controllers is the instability of the mean payoffs, measured by the deviations of the average rewards per transition computed in a finite "window" sliding along a run. Unfortunately, the problem of simultaneously optimizing the mean payoffs under local stability constraints is computationally hard, and the existing works do not provide a practically usable algorithm even for non-stochastic models such as two-player games. In this paper, we design and evaluate the first efficient and scalable solution to this problem applicable to Markov decision processes.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e5544c11-4eb3-4a1e-9208-118d0d7b7ee2Related papers
- PAC Statistical Model Checking of Mean Payoff in Discrete- and Continuous-Time MDPChaitanya Agarwal, Shibashis Guha, Jan Kretínský, Pazhamalai MuruganandhamCAV 2022 · 8 citations
- Optimizing Local Satisfaction of Long-Run Average Objectives in Markov Decision ProcessesDavid Klaska, Antonín Kucera, Vojtech Kur, Vít Musil et al.AAAI 2024 · 1 citation
- Stopping Criteria for Value Iteration on Stochastic Games with Quantitative ObjectivesJan Kretínský, Tobias Meggendorfer, Maximilian WeiningerLICS 2023 · 10 citations
- Variance Penalized On-Policy and Off-Policy Actor-CriticArushi Jain, Gandharv Patil, Ayush Jain, Khimya Khetarpal et al.AAAI 2021 · 11 citations
- Zero-Sum Games between Mean-Field Teams: Reachability-Based Analysis under Mean-Field SharingYue Guan, Mohammad Afshari, Panagiotis TsiotrasAAAI 2024 · 13 citations
