상세 보기
Multi-policy iteration with a distributed voting
Citations
WEB OF SCIENCE
3Citations
SCOPUS
3초록
We present a novel simulation-based algorithm, as an extension of the well-known policy iteration algorithm, by combining multi-policy improvement with a distributed simulation-based voting policy evaluation, for approximately solving Markov Decision Processes (MDPs) with infinite horizon discounted reward criterion, and analyze its performance relative to the optimal value.
키워드
policy iteration; distributed algorithm; voting; Markov decision processes
- 제목
- Multi-policy iteration with a distributed voting
- 저자
- Chang, HS
- 발행일
- 2004-11
- 유형
- Article
- 권
- 60
- 호
- 2
- 페이지
- 299 ~ 310