Value set iteration for two-person zero-sum Markov games

Citations

WEB OF SCIENCE

1
Citations

SCOPUS

1

초록

We present a novel exact algorithm called "value set iteration" (VSI) for solving two-person zero-sum Markov games (MGs) as a generalization of value iteration (VI) and as a general framework of combining multiple solution methods. We introduce a novel operator in the value function space and iteratively apply the operator with any sequence of the set of policies, extending Chang's VSI for MDPs into the MG setting. We show that VSI for MGs converges to the equilibrium value function with at least linear convergence rate and establish that VSI can potentially improve the convergence speed in terms of the number of iterations by proper setting of the sequence of the set of policies. (C) 2016 Elsevier Ltd. All rights reserved.

키워드

Two-person zero-sum Markov gameValue iterationPolicy iterationStochastic game
제목
Value set iteration for two-person zero-sum Markov games
저자
Chang, Hyeong Soo
DOI
10.1016/j.automatica.2016.10.010
발행일
2017-02
유형
Article
저널명
Automatica
76
페이지
61 ~ 64