Proximal-proximal-gradient method

11Citations
Citations of this article
84Readers
Mendeley users who have this article in their library.

Abstract

In this paper, we present the proximal-proximal-gradient method (PPG), a novel optimization method that is simple to implement and simple to parallelize. PPG generalizes the proximal-gradient method and ADMM and is applicable to minimization problems written as a sum of many differentiable and many non-differentiable convex functions. The non-differentiable functions can be coupled. We furthermore present a related stochastic variation, which we call stochastic PPG (S-PPG). S-PPG can be interpreted as a generalization of Finito and MISO over to the sum of many coupled non-differentiable convex functions. We present many applications that can benefit from PPG and S-PPG and prove convergence for both methods. We demonstrate the empirical effectiveness of both methods through experiments on a CUDA GPU. A key strength of PPG and S-PPG is, compared to existing methods, their ability to directly handle a large sum of non-differentiable nonseparable functions with a constant stepsize independent of the number of functions. Such non-diminishing stepsizes allows them to be fast.

Cite

CITATION STYLE

APA

Ryu, E. K., & Yin, W. (2019). Proximal-proximal-gradient method. Journal of Computational Mathematics, 37(6), 778–812. https://doi.org/10.4208/jcm.1906-m2018-0282

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free