An exploration strategy for non-stationary opponents-Reference-Cited by-同舟云学术

An exploration strategy for non-stationary opponents

Published:2016-10-13 Issue:5 Volume:31 Page:971-1002
ISSN:1387-2532
Container-title:Autonomous Agents and Multi-Agent Systems
language:en
Short-container-title:Auton Agent Multi-Agent Syst

Author:

Hernandez-Leal Pablo,Zhan Yusen,Taylor Matthew E.,Sucar L. Enrique,Munoz de Cote Enrique

Publisher

Springer Science and Business Media LLC

Subject

Artificial Intelligence

Link

http://link.springer.com/article/10.1007/s10458-016-9347-3/fulltext.html

Reference52 articles.

1. Auer, P., Cesa-Bianchi, N., & Fischer, P. (2002). Finite-time analysis of the multiarmed Bandit problem. Machine Learning, 47(2/3), 235–256.

2. Axelrod, R., & Hamilton, W. D. (1981). The evolution of cooperation. Science, 211(27), 1390–1396.

3. Babes, M., Munoz de Cote, E., & Littman, M. L. (2008). Social reward shaping in the prisoner’s dilemma. In Proceedings of the 7th International Conference on Autonomous Agents and Multiagent Systems, (pp. 1389–1392). Estoril: International Foundation for Autonomous Agents and Multiagent Systems.

4. Banerjee, B., & Peng, J. (2005). Efficient learning of multi-step best response. In Proceedings of the 4th International Conference on Autonomous Agents and Multiagent Systems, (pp. 60–66). Utretch: ACM.

5. Bard, N., Johanson, M., Burch, N., & Bowling, M. (2013). Online implicit agent modelling. In Proceedings of the 12th International Conference on Autonomous Agents and Multiagent Systems, (pp. 255–262). Saint Paul, MN: International Foundation for Autonomous Agents and Multiagent Systems.

Cited by 6 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Dual regularized policy updating and shiftpoint detection for automated deployment of reinforcement learning controllers on industrial mechatronic systems;Control Engineering Practice;2024-01

2. Defeating the Non-stationary Opponent Using Deep Reinforcement Learning and Opponent Modeling;Lecture Notes of the Institute for Computer Sciences, Social Informatics and Telecommunications Engineering;2024

3. An online learning algorithm to play discounted repeated games in wireless networks;Engineering Applications of Artificial Intelligence;2022-01

4. A kernel based learning method for non-stationary two-player repeated games;Knowledge-Based Systems;2020-05

5. Toll-based reinforcement learning for efficient equilibria in route choice;The Knowledge Engineering Review;2020