A stochastic derivative-free optimization method with importance sampling: Theory and learning to control

6Citations
Citations of this article
18Readers
Mendeley users who have this article in their library.

Abstract

We consider the problem of unconstrained minimization of a smooth objective function in Rn in a setting where only function evaluations are possible. While importance sampling is one of the most popular techniques used by machine learning practitioners to accelerate the convergence of their models when applicable, there is not much existing theory for this acceleration in the derivative-free setting. In this paper, we propose the first derivative free optimization method with importance sampling and derive new improved complexity results on non-convex, convex and strongly convex functions. We conduct extensive experiments on various synthetic and real LIBSVM datasets confirming our theoretical results. We test our method on a collection of continuous control tasks on MuJoCo environments with varying difficulty. Experiments show that our algorithm is practical for high dimensional continuous control problems where importance sampling results in a significant sample complexity improvement.

Cite

CITATION STYLE

APA

Bibi, A., Bergou, E. H., Sener, O., Ghanem, B., & Richtárik, P. (2020). A stochastic derivative-free optimization method with importance sampling: Theory and learning to control. In AAAI 2020 - 34th AAAI Conference on Artificial Intelligence (pp. 3275–3282). AAAI press. https://doi.org/10.1609/aaai.v34i04.5727

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free