Pixel-Wise Grasp Detection via Twin Deconvolution and Multi-Dimensional Attention

26Citations
Citations of this article
2Readers
Mendeley users who have this article in their library.
Get full text

Abstract

The grasp detection is crucial to high-quality robotic grasping. Typically, the mainstream encoder-decoder regression solution is attractive due to its high accuracy and efficiency, however, it is still challenging to solve the checkerboard artifacts from the uneven overlap of convolution results in decoder, and features from the encoder also need to be further refined. In this paper, a novel pixel-wise grasp detection network is proposed, which is composed of an encoder, a multi-dimensional attention bottleneck, and a decoder based on twin deconvolution. The proposed decoder introduces a twin branch upon the original transposed convolution branch. Through the overlap degree matrix provided by the twin branch, the original branch is re-weighted and then the checkerboard artifacts of the original branch are eliminated. Besides, to deeply explore the intrinsic relationship of features and strengthen feature discrimination, residual multi-head self-attention, cross-amplitude attention, and channel attention are integrated together. As a result, adaptive feature refinement is achieved. The effectiveness of the proposed method is verified by experiments.

Cite

CITATION STYLE

APA

Ren, G., Geng, W., Guan, P., Cao, Z., & Yu, J. (2023). Pixel-Wise Grasp Detection via Twin Deconvolution and Multi-Dimensional Attention. IEEE Transactions on Circuits and Systems for Video Technology, 33(8), 4002–4010. https://doi.org/10.1109/TCSVT.2023.3237866

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free