Achieving Near-Optimal Individual Regret & Low Communications in Multi-Agent Bandits

Achieving Near-Optimal Individual Regret & Low Communications in Multi-Agent Bandits
复制标题

DOI:
--
复制
发表时间:
2023
期刊:
--
影响因子:
--
通讯作者:
Xuchuang Wang;L. Yang;Y. Chen;Xutong Liu;M. Hajiesmaili;D. Towsley;John C.S. Lui
Xuchuang Wang;L. Yang;Y. Chen;Xutong Liu;M. Hajiesmaili;D. Towsley;John C.S. Lui
中科院分区:
其他
文献类型:
--
作者:
Xuchuang Wang;L. Yang;Y. Chen;Xutong Liu;M. Hajiesmaili;D. Towsley;John C.S. Lui

文献摘要

相似文献

合作多智能体多武装土匪(CMA 2B)研究如何分布式代理合作发挥相同的多武装土匪游戏。现有的大多数CMA 2B工作都集中在最大化所有代理的群体绩效-所有代理的个人绩效的累积(即
Cooperative multi-agent multi-armed bandits ( CMA2B ) study how distributed agents cooperatively play the same multi-armed bandit game. Most existing CMA2B works focused on maximizing the group performance of all agents—the accumulation of all agents’ individual performance (i