Qualitative Controller Synthesis for Consumption Markov Decision Processes
Frantisek Blahoudek, Tomás Brázdil, Petr Novotný, Melkior Ornik, Pranay Thangeda, Ufuk Topcu
摘要
Consumption Markov Decision Processes (CMDPs) are probabilistic decision-making models of resource-constrained systems. In a CMDP, the controller possesses a certain amount of a critical resource, such as electric power. Each action of the controller can consume some amount of the resource. Resource replenishment is only possible in special reload states, in which the resource level can be reloaded up to the full capacity of the system. The task of the controller is to prevent resource exhaustion, i.e. ensure that the available amount of the resource stays non-negative, while ensuring an additional linear-time property. We study the complexity of strategy synthesis in consumption MDPs with almost-sure Büchi objectives. We show that the problem can be solved in polynomial time. We implement our algorithm and show that it can efficiently solve CMDPs modelling real-world scenarios.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Density Constrained Reinforcement LearningZengyi Qin, Yuxiao Chen, Chuchu FanICML 2021 · 被引用 40 次
- Threshold UCT: Cost-Constrained Monte Carlo Tree Search with Pareto CurvesMartin Kurecka, Václav Nevyhostený, Petr Novotný, Vít UncovskýAAAI 2025 · 被引用 1 次
- Optimization of Multi-Agent Flying Sidekick Traveling Salesman Problem over Road NetworksRuixiao Yang, Chuchu FanAAAI 2026
相关 Paper
- Achieving Õ(1/ε) Sample Complexity for Constrained Markov Decision ProcessJiashuo Jiang, Yinyu YeNeurIPS 2024 · 被引用 3 次
- A Sample-Efficient Algorithm for Episodic Finite-Horizon MDP with ConstraintsKrishna Chaitanya Kalagarla, Rahul Jain, Pierluigi NuzzoAAAI 2021 · 被引用 58 次
- Polynomial-Time Approximability of Constrained Reinforcement LearningJeremy McMahanICML 2025
- Fuel in Markov Decision Processes (FiMDP): A Practical Approach to ConsumptionFrantisek Blahoudek, Murat Cubuktepe, Petr Novotný, Melkior Ornik 等FM 2021 · 被引用 2 次
- Learning with Safety Constraints: Sample Complexity of Reinforcement Learning for Constrained MDPsAria HasanzadeZonuzy, Archana Bura, Dileep M. Kalathil, Srinivas ShakkottaiAAAI 2021 · 被引用 46 次
