Grasping Control of a Vision Robot Based on a Deep Attentive Deterministic Policy Gradient

5Citations
Citations of this article
13Readers
Mendeley users who have this article in their library.

This article is free to access.

Abstract

Reinforcement learning can achieve excellent performance in the field of robotic grasping if the grasping target is stable. However, during applications in the real world, robot needs to overcome the effects of a complex working environment with different types of target objects, so it is more difficult to maintain the quality of action planning, even in the same scene. In order to make an agent have the ability to plan actions in a more adaptive way, the deep attentive deterministic policy gradient algorithm is applied in this article. An attention region proposal network is used to select the message of the pre-exploration area. Then this message is calculated using the adaptive exploration method to regulate the strategy as the target changes. Furthermore, a stratified reward function, which is used to reduce the negative influence of miscellaneous information brought by the sparse reward matrix, is defined according to the distance between the end effector and the center of the pre-exploration area. The results show that the DADPG is able to produce a robust strategy with noise interference, and can train in a more efficient way due to the hierarchical reward function.

Cite

CITATION STYLE

APA

Ji, X., Xiong, F., Kong, W., Wei, D., & Shen, Z. (2022). Grasping Control of a Vision Robot Based on a Deep Attentive Deterministic Policy Gradient. IEEE Access, 10, 867–878. https://doi.org/10.1109/ACCESS.2021.3137821

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free