Pixel-Reasoning-Based Robotics Fine Grasping for Novel Objects with Deep EDINet Structure

Chaoquan Shi; Chunxiao Miao; Xungao Zhong; Xunyu Zhong; Huosheng Hu; Qiang Liu

doi:10.3390/s22114283

Pixel-Reasoning-Based Robotics Fine Grasping for Novel Objects with Deep EDINet Structure

Sensors (Basel). 2022 Jun 4;22(11):4283. doi: 10.3390/s22114283.

Authors

Chaoquan Shi¹, Chunxiao Miao¹, Xungao Zhong¹, Xunyu Zhong², Huosheng Hu³, Qiang Liu⁴

Affiliations

¹ School of Electrical Enginnering and Automation, Xiamen University of Technology, Xiamen 361024, China.
² School of Aerospace Engineering, Xiamen University, Xiamen 361005, China.
³ School of Computer Science and Electronic Engineering, University of Essex, Colchester CO4 3SQ, UK.
⁴ Department of Psychiatry, University of Oxford, Oxford OX1 2JD, UK.

Abstract

Robotics grasp detection has mostly used the extraction of candidate grasping rectangles; those discrete sampling methods are time-consuming and may ignore the potential best grasp synthesis. This paper proposes a new pixel-level grasping detection method on RGB-D images. Firstly, a fine grasping representation is introduced to generate the gripper configurations of parallel-jaw, which can effectively resolve the gripper approaching conflicts and improve the applicability to unknown objects in cluttered scenarios. Besides, the adaptive grasping width is used to adaptively represent the grasping attribute, which is fine for objects. Then, the encoder-decoder-inception convolution neural network (EDINet) is proposed to predict the fine grasping configuration. In our findings, EDINet uses encoder, decoder, and inception modules to improve the speed and robustness of pixel-level grasping detection. The proposed EDINet structure was evaluated on the Cornell and Jacquard dataset; our method achieves 98.9% and 96.1% test accuracy, respectively. Finally, we carried out the grasping experiment on the unknown objects, and the results show that the average success rate of our network model is 97.2% in a single object scene and 93.7% in a cluttered scene, which out-performs the state-of-the-art algorithms. In addition, EDINet completes a grasp detection pipeline within only 25 ms.

Keywords: EDINet deep network; pixel-level reasoning; robotics fine grasping.

MeSH terms

Hand Strength*
Neural Networks, Computer
Robotics* / methods

Abstract

MeSH terms

Grants and funding