arXiv · 1002.4862
Less Regret via Online Conditioning
Abstract
We analyze and evaluate an online gradient descent algorithm with adaptive per-coordinate adjustment of learning rates. Our algorithm can be thought of as an online version of batch gradient descent with a diagonal preconditioner. This approach leads to regret bounds that are stronger than those of standard online gradient descent for general online convex optimization problems. Experimentally, we show that our algorithm is competitive with state-of-the-art algorithms for large scale machine learning problems.
Explore related subjects
Keep this discovery
Matthew Streeter, H. Brendan McMahan. 2010-02-25. Less Regret via Online Conditioning. https://arxiv.org/abs/1002.4862
Cite the original work for its findings. Save a collection to share your selection of sources.