Fuzzy Baselines to Stabilize Policy Gradient Reinforcement Learning-Reference-Cited by-同舟云学术

Fuzzy Baselines to Stabilize Policy Gradient Reinforcement Learning

Published:2021-07-28 Issue: Volume: Page:436-446
ISSN:2367-3370
Container-title:Explainable AI and Other Applications of Fuzzy Techniques
language:
Short-container-title:

Author:

Surita Gabriela,Lemos Andre,Gomide Fernando

Publisher

Springer International Publishing

Link

https://link.springer.com/content/pdf/10.1007/978-3-030-82099-2_39

Reference26 articles.

1. Berenji, H.: Fuzzy q-learning: a new approach for fuzzy dynamic programming. In: Proceedings of 1994 IEEE 3rd International Fuzzy Systems Conference, vol. 1, pp. 486–491 (1994). https://doi.org/10.1109/FUZZY.1994.343737

2. Botvinick, M., Ritter, S., Wang, J.X., Kurth-Nelson, Z., Blundell, C., Hassabis, D.: Reinforcement learning, fast and slow. Trends Cogn. Sci. 23(5), 408–422 (2019)

3. Brockman, G., et al.: Openai gym. arXiv preprint arXiv:1606.01540 (2016)

4. Cooper, M.G., Vidal, J.J.: Genetic design of fuzzy controllers: the cart and jointed-pole problem. In: Proceedings of 1994 IEEE 3rd International Fuzzy Systems Conference, pp. 1332–1337. IEEE (1994)

5. Duan, Y., et al.: One-shot imitation learning. In: Guyon, I., et al. (eds.) Advances in Neural Information Processing Systems, vol. 30, pp. 1087–1098. Curran Associates, Inc. (2017). http://papers.nips.cc/paper/6709-one-shot-imitation-learning.pdf

Cited by 1 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Function approximation reinforcement learning of energy management with the fuzzy REINFORCE for fuel cell hybrid electric vehicles;Energy and AI;2023-07