An Approximate Quadratic Programming for Efficient Bellman Equation Solution

5Citations
Citations of this article
8Readers
Mendeley users who have this article in their library.
Get full text

Abstract

This paper proposes an efficient algorithm which relies on quadratic programming for approximately solving the Bellman equation in reinforcement learning problem and guarantees to return optimal decision parameters. Through further applying universal approximation and fixed cardinality minimization techniques, the proposed algorithm in one hand expands the representation ability of basic linear value functions, on the other hand, it guarantees the convergence of the Bellman error. Experimental results on two canonical reinforcement learning scenarios demonstrate that the proposed algorithm achieves similar or better performance than the state-of-The-Art algorithms, while reduces the computation time significantly and improves the robustness of the algorithm against state uncertainty.

Cite

CITATION STYLE

APA

Su, J., Cheng, H., Guo, H., Huang, R., & Peng, Z. (2019). An Approximate Quadratic Programming for Efficient Bellman Equation Solution. IEEE Access, 7, 126077–126087. https://doi.org/10.1109/ACCESS.2019.2939161

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free