Achieving Socially Optimal Outcomes in Multiagent Systems with Reinforcement Social Learning-Reference-Cited by-同舟云学术

Achieving Socially Optimal Outcomes in Multiagent Systems with Reinforcement Social Learning

Published:2013-09 Issue:3 Volume:8 Page:1-23
ISSN:1556-4665
Container-title:ACM Transactions on Autonomous and Adaptive Systems
language:en
Short-container-title:ACM Trans. Auton. Adapt. Syst.

Author:

Hao Jianye¹,Leung Ho-Fung¹

Affiliation:

1. The Chinese University of Hong Kong

Abstract

In multiagent systems, social optimality is a desirable goal to achieve in terms of maximizing the global efficiency of the system. We study the problem of coordinating on socially optimal outcomes among a population of agents, in which each agent randomly interacts with another agent from the population each round. Previous work [Hales and Edmonds 2003; Matlock and Sen 2007, 2009] mainly resorts to modifying the interaction protocol from random interaction to tag-based interactions and only focus on the case of symmetric games. Besides, in previous work the agents’ decision making processes are usually based on evolutionary learning, which usually results in high communication cost and high deviation on the coordination rate. To solve these problems, we propose an alternative social learning framework with two major contributions as follows. First, we introduce the observation mechanism to reduce the amount of communication required among agents. Second, we propose that the agents’ learning strategies should be based on reinforcement learning technique instead of evolutionary learning. Each agent explicitly keeps the record of its current state in its learning strategy, and learn its optimal policy for each state independently. In this way, the learning performance is much more stable and also it is suitable for both symmetric and asymmetric games. The performance of this social learning framework is extensively evaluated under the testbed of two-player general-sum games comparing with previous work [Hao and Leung 2011; Matlock and Sen 2007]. The influences of different factors on the learning performance of the social learning framework are investigated as well.

Publisher

Association for Computing Machinery (ACM)

Subject

Software,Computer Science (miscellaneous),Control and Systems Engineering

Link

https://dl.acm.org/doi/pdf/10.1145/2517329

Reference35 articles.

1. Allison P. D. 1992. The cultural evolution of beneficent norms. Social Forces. 279--301. Allison P. D. 1992. The cultural evolution of beneficent norms. Social Forces . 279--301.

2. Multiagent learning using a variable learning rate

3. Efficient learning equilibrium

4. Tag Mechanisms Evaluated for Coordination in Open Multi-Agent Systems

Cited by 16 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. The double-edged sword effect of conformity on cooperation in spatial Prisoner’s Dilemma Games with reinforcement learning;Chaos, Solitons & Fractals;2024-10

2. Learning in Cooperative Multiagent Systems Using Cognitive and Machine Models;ACM Transactions on Autonomous and Adaptive Systems;2023-10-14

3. How committed individuals shape social dynamics: A survey on coordination games and social dilemma games;Europhysics Letters;2023-10-01

4. Incorporating social payoff into reinforcement learning promotes cooperation;Chaos: An Interdisciplinary Journal of Nonlinear Science;2022-12

5. A literature review on optimization techniques for adaptation planning in adaptive systems: State of the art and research directions;Information and Software Technology;2022-09