Transition Based Discount Factor for Model Free Algorithms in Reinforcement Learning-Reference-Cited by-同舟云学术

Transition Based Discount Factor for Model Free Algorithms in Reinforcement Learning

Published:2021-07-02 Issue:7 Volume:13 Page:1197
ISSN:2073-8994
Container-title:Symmetry
language:en
Short-container-title:Symmetry

Author:

Sharma Abhinav,Gupta Ruchir^ORCID,Lakshmanan K.,Gupta Atul

Abstract

Reinforcement Learning (RL) enables an agent to learn control policies for achieving its long-term goals. One key parameter of RL algorithms is a discount factor that scales down future cost in the state’s current value estimate. This study introduces and analyses a transition-based discount factor in two model-free reinforcement learning algorithms: Q-learning and SARSA, and shows their convergence using the theory of stochastic approximation for finite state and action spaces. This causes an asymmetric discounting, favouring some transitions over others, which allows (1) faster convergence than constant discount factor variant of these algorithms, which is demonstrated by experiments on the Taxi domain and MountainCar environments; (2) provides better control over the RL agents to learn risk-averse or risk-taking policy, as demonstrated in a Cliff Walking experiment.

Publisher

MDPI AG

Subject

Physics and Astronomy (miscellaneous),General Mathematics,Chemistry (miscellaneous),Computer Science (miscellaneous)

Link

https://www.mdpi.com/2073-8994/13/7/1197/pdf

Reference41 articles.

1. Reinforcement Learning: An Introduction;Sutton,1998

2. Deep reinforcement learning optimization framework for a power generation plant considering performance and environmental issues

3. Grandmaster level in StarCraft II using multi-agent reinforcement learning

4. Testing match-3 video games with Deep Reinforcement Learning;Napolitano;arXiv,2020

Cited by 2 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Reinforcement Learning for Laser Welding Speed Control Minimizing Bead Width Error;2023 IEEE International Conference on Robotics and Automation (ICRA);2023-05-29

2. Towards the design of vision-based intelligent vehicle system: methodologies and challenges;Evolutionary Intelligence;2022-03-09