Recursive Markov Decision Processes and Recursive Stochastic Games

被引：13

作者：

Etessami, Kousha ^{[1
]}

Yannakakis, Mihalis ^{[2
]}

机构：

[1] Univ Edinburgh, Sch Informat, LFCS, Edinburgh EH8 9YL, Midlothian, Scotland

[2] Columbia Univ, Dept Comp Sci, New York, NY 10027 USA

来源：

JOURNAL OF THE ACM | 2015年 / 62卷 / 02期

基金：

美国国家科学基金会;

关键词：

Algorithms; Theory; Verification; Recursive stochastic processes; Markov decision processes; stochastic games; multitype branching processes; stochastic context-free grammars; MODEL CHECKING; MONOTONE SYSTEMS; COMPLEXITY; REACHABILITY; OPTIMIZATION; GRAMMARS; TIME;

D O I：

10.1145/2699431

中图分类号：

TP3 [计算技术、计算机技术];

学科分类号：

0812 ;

摘要：

We introduce Recursive Markov Decision Processes (RMDPs) and Recursive Simple Stochastic Games (RSSGs), which are classes of (finitely presented) countable-state MDPs and zero-sum turn-based (perfect information) stochastic games. They extend standard finite-state MDPs and stochastic games with a recursion feature. We study the decidability and computational complexity of these games under termination objectives for the two players: one player's goal is to maximize the probability of termination at a given exit, while the other player's goal is to minimize this probability. In the quantitative termination problems, given an RMDP (or RSSG) and probability p, we wish to decide whether the value of such a termination game is at least p (or at most p); in the qualitative termination problem we wish to decide whether the value is 1. The important 1-exit subclasses of these models, 1-RMDPs and 1-RSSGs, correspond in a precise sense to controlled and game versions of classic stochastic models, including multitype Branching Processes and Stochastic Context-Free Grammars, where the objective of the players is to maximize or minimize the probability of termination (extinction). We provide a number of upper and lower bounds for qualitative and quantitative termination problems for RMDPs and RSSGs. We show both problems are undecidable for multi-exit RMDPs, but are decidable for 1-RMDPs and 1-RSSGs. Specifically, the quantitative termination problem is decidable in PSPACE for both 1-RMDPs and 1-RSSGs, and is at least as hard as the square root sum problem, a well-known open problem in numerical computation. We show that the qualitative termination problem for 1-RMDPs (i.e., a controlled version of branching processes) can be solved in polynomial time both for maximizing and minimizing 1-RMDPs. The qualitative problem for 1-RSSGs is in NP boolean AND coNP, and is at least as hard as the quantitative termination problem for Condon's finite-state simple stochastic games, whose complexity remains a well known open problem. Finally, we show that even for 1-RMDPs, more general (qualitative and quantitative) model-checking problems with respect to linear-time temporal properties are undecidable even for a fixed property.

引用

页数：69

共 50 条

[1] Recursive Markov decision processes and recursive stochastic games
Etessami, K
Yannakakis, M
[J]. AUTOMATA, LANGUAGES AND PROGRAMMING, PROCEEDINGS, 2005, 3580 : 891 - 903
[2] Efficient qualitative analysis of classes of recursive Markov decision processes and simple Stochastic games
Etessami, K
Yannakakis, M
[J]. STACS 2006, PROCEEDINGS, 2006, 3884 : 634 - 645
[3] Reachability in recursive Markov decision processes
Brazdil, Tomas
Brozek, Vaclav
Forejt, Vojtech
Kucera, Antonin
[J]. CONCUR 2006 - CONCURRENCY THEORY, PROCEEDINGS, 2006, 4137 : 358 - 374
[4] Reachability in recursive Markov decision processes
Brazdil, Tomas
Brozek, Vaclav
Forejt, Vojtech
Kucera, Antonin
[J]. INFORMATION AND COMPUTATION, 2008, 206 (05) : 520 - 537
[5] Markov decision processes with recursive risk measures
Baeuerle, Nicole
Glauner, Alexander
[J]. EUROPEAN JOURNAL OF OPERATIONAL RESEARCH, 2022, 296 (03) : 953 - 966
[6] Recursive learning automata approach to Markov decision processes
Chang, Hyeong Soo
Fu, Michael C.
Hu, Jiaqiao
Marcus, Steven I.
[J]. IEEE TRANSACTIONS ON AUTOMATIC CONTROL, 2007, 52 (07) : 1349 - 1355
[7] RECURSIVE CONCURRENT STOCHASTIC GAMES
Etessami, Kousha
Yannakakis, Mihalis
[J]. LOGICAL METHODS IN COMPUTER SCIENCE, 2008, 4 (04)
[8] Recursive Concurrent Stochastic Games
Etessami, Kousha
Yannakakis, Mihalis
[J]. AUTOMATA, LANGUAGES AND PROGRAMMING, PT 2, 2006, 4052 : 324 - 335
[9] Reducible Markov Decision Processes and Stochastic Games
Ning, Jie
[J]. PRODUCTION AND OPERATIONS MANAGEMENT, 2021, 30 (08) : 2726 - 2751
[10] Recursive stochastic games with positive rewards
Etessami, Kousha
Wojtczak, Dominik
Yannakakis, Mihalis
[J]. THEORETICAL COMPUTER SCIENCE, 2019, 777 : 308 - 328

← 1 2 3 4 5 →