One-shot sim-to-real transfer policy for robotic assembly via reinforcement learning with visual demonstration-Reference-Cited by-同舟云学术

One-shot sim-to-real transfer policy for robotic assembly via reinforcement learning with visual demonstration

Published:2024-01-24 Issue: Volume: Page:1-20
ISSN:0263-5747
Container-title:Robotica
language:en
Short-container-title:Robotica

Author:

Xiao Ruihong^ORCID,Yang Chenguang^ORCID,Jiang Yiming,Zhang Hui

Abstract

Abstract Reinforcement learning (RL) has been successfully applied to a wealth of robot manipulation tasks and continuous control problems. However, it is still limited to industrial applications and suffers from three major challenges: sample inefficiency, real data collection, and the gap between simulator and reality. In this paper, we focus on the practical application of RL for robot assembly in the real world. We apply enlightenment learning to improve the proximal policy optimization, an on-policy model-free actor-critic reinforcement learning algorithm, to train an agent in Cartesian space using the proprioceptive information. We introduce enlightenment learning incorporated via pretraining, which is beneficial to reduce the cost of policy training and improve the effectiveness of the policy. A human-like assembly trajectory is generated through a two-step method with segmenting objects by locations and iterative closest point for pretraining. We also design a sim-to-real controller to correct the error while transferring to reality. We set up the environment in the MuJoCo simulator and demonstrated the proposed method on the recently established The National Institute of Standards and Technology (NIST) gear assembly benchmark. The paper introduces a unique framework that enables a robot to learn assembly tasks efficiently using limited real-world samples by leveraging simulations and visual demonstrations. The comparative experiment results indicate that our approach surpasses other baseline methods in terms of training speed, success rate, and efficiency.

Publisher

Cambridge University Press (CUP)

Subject

Computer Science Applications,General Mathematics,Software,Control and Systems Engineering,Control and Optimization,Mechanical Engineering,Modeling and Simulation,Artificial Intelligence,Computer Vision and Pattern Recognition,Computational Mechanics,Rehabilitation

Reference49 articles.

1. Path finding and grasp planning for robotic assembly;Lee;Robotica,1994

2. [36] He, K. , Gkioxari, G. , Dollár, P. and Girshick, R. , “Mask R-CNN” Proceedings of the IEEE International Conference on Computer Vision (2017) pp. 2961–2969.

3. [38] Zakharov, S. , Shugurov, I. and Ilic, S. , “Dpod: 6D Pose Object Detector and Refiner” Proceedings of the IEEE/CVF International Conference on Computer Vision (2019) pp. 1941–1950.

4. Efficient experience replay architecture for offline reinforcement learning;Zhang;Robot. Intell. Automat.,2023

5. [44] Arndt, K. , Hazara, M. , Ghadirzadeh, A. and Kyrki, V. , “Meta Reinforcement Learning for Sim-to-Real Domain Adaptation” 2020 IEEE International Conference on Robotics and Automation (ICRA), IEEE (2020) pp. 2725–2731.