LTL and Beyond: Formal Languages for Reward Function Specification in Reinforcement Learning-Reference-Cited by-同舟云学术

LTL and Beyond: Formal Languages for Reward Function Specification in Reinforcement Learning

Published:2019-08 Issue: Volume: Page:
ISSN:
Container-title:Proceedings of the Twenty-Eighth International Joint Conference on Artificial Intelligence
language:
Short-container-title:

Author:

Camacho Alberto¹²,Toro Icarte Rodrigo¹²,Klassen Toryn Q.¹,Valenzano Richard³,McIlraith Sheila A.¹²

Affiliation:

1. Department of Computer Science, University of Toronto, Toronto, Canada

2. Vector Institute, Toronto, Canada

3. Element AI, Toronto, Canada

Abstract

In Reinforcement Learning (RL), an agent is guided by the rewards it receives from the reward function. Unfortunately, it may take many interactions with the environment to learn from sparse rewards, and it can be challenging to specify reward functions that reflect complex reward-worthy behavior. We propose using reward machines (RMs), which are automata-based representations that expose reward function structure, as a normal form representation for reward functions. We show how specifications of reward in various formal languages, including LTL and other regular languages, can be automatically translated into RMs, easing the burden of complex reward function specification. We then show how the exposed structure of the reward function can be exploited by tailored q-learning algorithms and automated reward shaping techniques in order to improve the sample efficiency of reinforcement learning methods. Experiments show that these RM-tailored techniques significantly outperform state-of-the-art (deep) RL algorithms, solving problems that otherwise cannot reasonably be solved by existing approaches.

Publisher

International Joint Conferences on Artificial Intelligence Organization

Cited by 49 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Generalization of temporal logic tasks via future dependent options;Machine Learning;2024-08-26

2. Decomposition-Based Hierarchical Task Allocation and Planning for Multi-Robots Under Hierarchical Temporal Logic Specifications;IEEE Robotics and Automation Letters;2024-08

3. Keeping Behavioral Programs Alive: Specifying and Executing Liveness Requirements;2024 IEEE 32nd International Requirements Engineering Conference (RE);2024-06-24

4. A Review of Reward Functions for Reinforcement Learning in the context of Autonomous Driving;2024 IEEE Intelligent Vehicles Symposium (IV);2024-06-02

5. Runtime Verification-Based Safe MARL for Optimized Safety Policy Generation for Multi-Robot Systems;Big Data and Cognitive Computing;2024-05-16