Asynchronous Upper Confidence Bound Algorithms for Federated Linear Bandits
Asynchronous Upper Confidence Bound Algorithms for Federated Linear Bandits
复制标题
DOI:
--
复制
发表时间:
2021-10
期刊:
影响因子:
--
通讯作者:
Chuanhao Li;Hongning Wang
中科院分区:
文献类型:
--
作者:
Chuanhao Li;Hongning Wang
Linear contextual bandit is a popular online learning problem. It has been mostly studied in centralized learning settings. With the surging demand of large-scale decentralized model learning, e.g., federated learning, how to retain regret minimization while reducing communication cost becomes an open challenge. In this paper, we study linear contextual bandit in a federated learning setting. We propose a general framework with asynchronous model update and communication for a collection of homogeneous clients and heterogeneous clients, respectively. Rigorous theoretical analysis is provided about the regret and communication cost under this distributed learning framework; and extensive empirical evaluations demonstrate the effectiveness of our solution.