← 返回论文检索
AAAI 2024official proceedings

Optimizing Local Satisfaction of Long-Run Average Objectives in Markov Decision Processes

David Klaška, Antonín Kučera, Vojtěch Kůr, Vít Musil, Vojtěch Řehák

PDF 由论文原始站点提供,PaperCompass 不保存论文文件。DOI 10.1609/aaai.v38i18.29993 ↗

摘要

Long-run average optimization problems for Markov decision processes (MDPs) require constructing policies with optimal steady-state behavior, i.e., optimal limit frequency of visits to the states. However, such policies may suffer from local instability in the sense that the frequency of states visited in a bounded time horizon along a run differs significantly from the limit frequency. In this work, we propose an efficient algorithmic solution to this problem.