Hybrid Imitation Learning Framework for Robotic Manipulation Tasks

Eunjin Jung; Incheol Kim

doi:10.3390/s21103409

Hybrid Imitation Learning Framework for Robotic Manipulation Tasks

Sensors (Basel). 2021 May 13;21(10):3409. doi: 10.3390/s21103409.

Authors

Eunjin Jung¹, Incheol Kim¹

Affiliation

¹ Department of Computer Science, Kyonggi University, Suwon-si 16227, Korea.

Abstract

This study proposes a novel hybrid imitation learning (HIL) framework in which behavior cloning (BC) and state cloning (SC) methods are combined in a mutually complementary manner to enhance the efficiency of robotic manipulation task learning. The proposed HIL framework efficiently combines BC and SC losses using an adaptive loss mixing method. It uses pretrained dynamics networks to enhance SC efficiency and performs stochastic state recovery to ensure stable learning of policy networks by transforming the learner's task state into a demo state on the demo task trajectory during SC. The training efficiency and policy flexibility of the proposed HIL framework are demonstrated in a series of experiments conducted to perform major robotic manipulation tasks (pick-up, pick-and-place, and stack tasks). In the experiments, the HIL framework showed about a 2.6 times higher performance improvement than the pure BC and about a four times faster training time than the pure SC imitation learning method. In addition, the HIL framework also showed about a 1.6 times higher performance improvement and about a 2.2 times faster training time than the other hybrid learning method combining BC and reinforcement learning (BC + RL) in the experiments.

Keywords: behavior cloning; dynamics modeling; hybrid imitation learning; robotic object manipulation task; trajectory cloning.

Grants and funding

2020-0-00096/Institute for Information & Communications Technology Planning & Evaluation (IITP) / Development of ICT Technology Program