Adaptive Restart for Accelerated Gradient Schemes

Brendan O'Donoghue; Emmanuel Candes

arXiv:1204.3982·math.OC·April 19, 2012·Found. Comput. Math.·21 cites

Adaptive Restart for Accelerated Gradient Schemes

Brendan O'Donoghue, Emmanuel Candes

PDF

Open Access

TL;DR

This paper introduces a simple heuristic adaptive restart method for accelerated gradient schemes that significantly enhances convergence rates by resetting momentum based on observed periodic behavior, often achieving optimal convergence without prior parameter knowledge.

Contribution

The paper proposes a novel adaptive restart heuristic for accelerated gradient methods, analyzing its effectiveness in improving convergence without needing prior knowledge of function parameters.

Findings

01

Adaptive restart improves convergence speed.

02

Periodic behavior indicates when to restart momentum.

03

Method often recovers optimal convergence rates.

Abstract

In this paper we demonstrate a simple heuristic adaptive restart technique that can dramatically improve the convergence rate of accelerated gradient schemes. The analysis of the technique relies on the observation that these schemes exhibit two modes of behavior depending on how much momentum is applied. In what we refer to as the 'high momentum' regime the iterates generated by an accelerated gradient scheme exhibit a periodic behavior, where the period is proportional to the square root of the local condition number of the objective function. This suggests a restart technique whereby we reset the momentum whenever we observe periodic behavior. We provide analysis to show that in many cases adaptively restarting allows us to recover the optimal rate of convergence with no prior knowledge of function parameters.

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsSparse and Compressive Sensing Techniques · Stochastic Gradient Optimization Techniques · Numerical methods in inverse problems