Learning from demonstration using products of experts: Applications to manipulation and task prioritization-Reference-Cited by-同舟云学术

Learning from demonstration using products of experts: Applications to manipulation and task prioritization

Published:2021-09-22 Issue:2 Volume:41 Page:163-188
ISSN:0278-3649
Container-title:The International Journal of Robotics Research
language:en
Short-container-title:The International Journal of Robotics Research

Author:

Pignat Emmanuel¹²,Silvério Joāo¹^ORCID,Calinon Sylvain¹²^ORCID

Affiliation:

1. Idiap Research Institute, Martigny, Switzerland

2. EPFL, Lausanne, Switzerland

Abstract

Probability distributions are key components of many learning from demonstration (LfD) approaches, with the spaces chosen to represent tasks playing a central role. Although the robot configuration is defined by its joint angles, end-effector poses are often best explained within several task spaces. In many approaches, distributions within relevant task spaces are learned independently and only combined at the control level. This simplification implies several problems that are addressed in this work. We show that the fusion of models in different task spaces can be expressed as products of experts (PoE), where the probabilities of the models are multiplied and renormalized so that it becomes a proper distribution of joint angles. Multiple experiments are presented to show that learning the different models jointly in the PoE framework significantly improves the quality of the final model. The proposed approach particularly stands out when the robot has to learn hierarchical objectives that arise when a task requires the prioritization of several sub-tasks (e.g. in a humanoid robot, keeping balance has a higher priority than reaching for an object). Since training the model jointly usually relies on contrastive divergence, which requires costly approximations that can affect performance, we propose an alternative strategy using variational inference and mixture model approximations. In particular, we show that the proposed approach can be extended to PoE with a nullspace structure (PoENS), where the model is able to recover secondary tasks that are masked by the resolution of tasks of higher-importance.

Publisher

SAGE Publications

Subject

Applied Mathematics,Artificial Intelligence,Electrical and Electronic Engineering,Mechanical Engineering,Modeling and Simulation,Software

Link

http://journals.sagepub.com/doi/pdf/10.1177/02783649211040561

Reference69 articles.

1. A randomized roadmap method for path and manipulation planning

2. Ergodic coverage in constrained environments using stochastic trajectory optimization

Cited by 11 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Learning and generalization of task-parameterized skills through few human demonstrations;Engineering Applications of Artificial Intelligence;2024-07

2. A Probabilistic Approach to Multi-Modal Adaptive Virtual Fixtures;IEEE Robotics and Automation Letters;2024-06

3. Sustainable manufacturing through application of reconfigurable and intelligent systems in production processes: a system perspective;Scientific Reports;2023-12-16

4. Tensor train for global optimization problems in robotics;The International Journal of Robotics Research;2023-11-30

5. Robust Control of Linear Systems: A Min-Max Reinforcement Learning Formulation;2023 20th International Conference on Electrical Engineering, Computing Science and Automatic Control (CCE);2023-10-25