Adaptive Optimization of Hyper-Parameters for Robotic Manipulation through Evolutionary Reinforcement Learning-Reference-Cited by-同舟云学术

Adaptive Optimization of Hyper-Parameters for Robotic Manipulation through Evolutionary Reinforcement Learning

Published:2024-07-24 Issue:3 Volume:110 Page:
ISSN:1573-0409
Container-title:Journal of Intelligent & Robotic Systems
language:en
Short-container-title:J Intell Robot Syst

Author:

Onori Giulio,Shahid Asad Ali,Braghin Francesco,Roveda Loris^ORCID

Abstract

AbstractDeep Reinforcement Learning applications are growing due to their capability of teaching the agent any task autonomously and generalizing the learning. However, this comes at the cost of a large number of samples and interactions with the environment. Moreover, the robustness of learned policies is usually achieved by a tedious tuning of hyper-parameters and reward functions. In order to address this issue, this paper proposes an evolutionary RL algorithm for the adaptive optimization of hyper-parameters. The policy is trained using an on-policy algorithm, Proximal Policy Optimization (PPO), coupled with an evolutionary algorithm. The achieved results demonstrate an improvement in the sample efficiency of the RL training on a robotic grasping task. In particular, the learning is improved with respect to the baseline case of a non-evolutionary agent. The evolutionary agent needs

$$60$$

60 % fewer samples to completely learn the grasping task, enabled by the adaptive transfer of knowledge between the agents through the evolutionary algorithm. The proposed approach also demonstrates the possibility of updating reward parameters during training, potentially providing a general approach to creating reward functions.

Funder

Hasler Stiftung

Publisher

Springer Science and Business Media LLC

Link

https://link.springer.com/content/pdf/10.1007/s10846-024-02138-8.pdf

Reference49 articles.

1. Bai, Q., Li, S., Yang, J., Song, Q., Li, Z., Zhang, X.: Object detection recognition and robot grasping based on machine learning: A survey. IEEE Access 8, 181855–181879 (2020)

2. Semeraro, F., Griffiths, A., Cangelosi, A.: Human–robot collaboration and machine learning: A systematic review of recent research. Robot. Comput.-Integrated Manufac. 79, 102432 (2023)

3. Song, X., Sun, P., Song, S., Stojanovic, V.: Quantized neural adaptive finite-time preassigned performance control for interconnected nonlinear systems. Neural Comput. Appl. 35(21), 15429–15446 (2023)

4. Tao, H., Zheng, J., Wei, J., Paszke, W., Rogers, E., Stojanovic, V.: Repetitive process based indirect-type iterative learning control for batch processes with model uncertainty and input delay. J. Process Control 132, 103112 (2023)

5. Billard, A.G., Calinon, S., Dillmann, R.: Learning from humans. Springer handbook of robotics, 1995–2014 (2016)