Combining multiple strategies for multiarmed bandit problems and asymptotic optimality

Citations

SCOPUS

0

초록

This brief paper provides a simple algorithm that selects a strategy at each time in a given set of multiple strategies for stochastic multiarmed bandit problems, thereby playing the arm by the chosen strategy at each time. The algorithm follows the idea of the probabilistic ε t -switching in the ε t -greedy strategy and is asymptotically optimal in the sense that the selected strategy converges to the best in the set under some conditions on the strategies in the set and the sequence of { ε t }. © 2015 Hyeong Soo Chang and Sanghee Choe.

제목
Combining multiple strategies for multiarmed bandit problems and asymptotic optimality
저자
Chang, Hyeong SooChoe, Sanghee
DOI
10.1155/2015/264953
발행일
2015
유형
Article
저널명
Journal of Control Science and Engineering
2015