Multi-Armed Bandits: Theory and Applications to Online Learning in Networks
Multi-Armed Bandits: Theory and Applications to Online Learning in Networks
复制标题
多臂强盗:网络在线学习的理论与应用
DOI:
10.2200/s00941ed2v01y201907cnt022
复制
发表时间:
2019
期刊:
影响因子:
--
通讯作者:
Qing Zhao
中科院分区:
文献类型:
--
作者:
Qing Zhao
Abstract Multi-armed bandit problems pertain to optimal sequential decision making and learning in unknown environments. Since the first bandit problem posed by Thompson in 1933 for the application...
影响因子:
1.5
作者:
Bull A
通讯作者:
Bull A
影响因子:
2.5
作者:
Kleinberg, Robert;Slivkins, Aleksandrs;Upfal, Eli
通讯作者:
Upfal, Eli