Online Non-Convex Learning: Following the Perturbed Leader is Optimal
Online Non-Convex Learning: Following the Perturbed Leader is Optimal
复制标题
DOI:
--
复制
发表时间:
2019-03
期刊:
影响因子:
--
通讯作者:
A. Suggala;Praneeth Netrapalli
中科院分区:
文献类型:
--
作者:
A. Suggala;Praneeth Netrapalli
We study the problem of online learning with non-convex losses, where the learner has access to an offline optimization oracle. We show that the classical Follow the Perturbed Leader (FTPL) algorithm achieves optimal regret rate of $O(T^{-1/2})$ in this setting. This improves upon the previous best-known regret rate of $O(T^{-1/3})$ for FTPL. We further show that an optimistic variant of FTPL achieves better regret bounds when the sequence of losses encountered by the learner is `predictable'.