상세 보기
A policy improvement method for constrained average Markov decision processes
Citations
WEB OF SCIENCE
8Citations
SCOPUS
10초록
This brief paper presents a policy improvement method for constrained Markov decision processes (MDPs) with average cost criterion under an ergodicity assumption, extending Howard's policy improvement for MDPs. The improvement method induces a policy iteration-type algorithm that converges to a local optimal policy. (c) 2006 Elsevier B.V. All rights reserved.
키워드
constrained Markov decision process; policy improvement; policy iteration
- 제목
- A policy improvement method for constrained average Markov decision processes
- 저자
- Chang, Hyeong Soo
- 발행일
- 2007-07
- 유형
- Article
- 권
- 35
- 호
- 4
- 페이지
- 434 ~ 438