Dual Averaging Methods for Regularized Stochastic Learning and Online Optimization
Dual Averaging Methods for Regularized Stochastic Learning and Online Optimization
复制标题
DOI:
10.5555/1756006.1953017
复制
发表时间:
2009-12
期刊:
影响因子:
--
通讯作者:
Lin Xiao
中科院分区:
文献类型:
--
作者:
Lin Xiao
We consider regularized stochastic learning and online optimization problems, where the objective function is the sum of two convex terms: one is the loss function of the learning task, and the other is a simple regularization term such as l1-norm for promoting sparsity. We develop a new online algorithm, the regularized dual averaging (RDA) method, that can explicitly exploit the regularization structure in an online setting. In particular, at each iteration, the learning variables are adjusted by solving a simple optimization problem that involves the running average of all past subgradients of the loss functions and the whole regularization term, not just its subgradient. Computational experiments show that the RDA method can be very effective for sparse online learning with l1-regularization.