Risk-Sensitive Reinforcement Learning Applied to Control under Constraints-Reference-Cited by-同舟云学术

Risk-Sensitive Reinforcement Learning Applied to Control under Constraints

Published:2005-07-01 Issue: Volume:24 Page:81-108
ISSN:1076-9757
Container-title:Journal of Artificial Intelligence Research
language:
Short-container-title:jair

Author:

Geibel P.,Wysotzki F.

Abstract

In this paper, we consider Markov Decision Processes (MDPs) with error states. Error states are those states entering which is undesirable or dangerous. We define the risk with respect to a policy as the probability of entering such a state when the policy is pursued. We consider the problem of finding good policies whose risk is smaller than some user-specified threshold, and formalize it as a constrained MDP with two criteria. The first criterion corresponds to the value function originally given. We will show that the risk can be formulated as a second criterion function based on a cumulative return, whose definition is independent of the original value function. We present a model free, heuristic reinforcement learning algorithm that aims at finding good deterministic policies. It is based on weighting the original value function and the risk. The weight parameter is adapted in order to find a feasible solution for the constrained problem that has a good performance with respect to the value function. The algorithm was successfully applied to the control of a feed tank with stochastic inflows that lies upstream of a distillation column. This control task was originally formulated as an optimal control problem with chance constraints, and it was solved under certain assumptions on the model to obtain an optimal solution. The power of our learning algorithm is that it can be used even when some of these restrictive assumptions are relaxed.

Publisher

AI Access Foundation

Subject

Artificial Intelligence

Cited by 142 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Sample-based bounds for coherent risk measures: Applications to policy synthesis and verification;Artificial Intelligence;2024-11

2. Reset-Free Reinforcement Learning via Multi-State Recovery and Failure Prevention for Autonomous Robots;Tsinghua Science and Technology;2024-10

3. Performance Evaluation of Fractional Proportional–Integral–Derivative Controllers Tuned by Heuristic Algorithms for Nonlinear Interconnected Tanks;Algorithms;2024-07-10

4. Towards Robust Decision-Making for Autonomous Highway Driving Based on Safe Reinforcement Learning;Sensors;2024-06-26

5. Synthesize Efficient Safety Certificates for Learning-Based Safe Control using Magnitude Regularization;2024 IEEE International Conference on Robotics and Automation (ICRA);2024-05-13