Mixture of personality improved spiking actor network for efficient multi-agent cooperation-Reference-Cited by-同舟云学术

Mixture of personality improved spiking actor network for efficient multi-agent cooperation

Published:2023-07-06 Issue: Volume:17 Page:
ISSN:1662-453X
Container-title:Frontiers in Neuroscience
language:
Short-container-title:Front. Neurosci.

Author:

Li Xiyun,Ni Ziyi,Ruan Jingqing,Meng Linghui,Shi Jing,Zhang Tielin,Xu Bo

Abstract

Adaptive multi-agent cooperation with especially unseen partners is becoming more challenging in multi-agent reinforcement learning (MARL) research, whereby conventional deep-learning-based algorithms suffer from the poor new-player-generalization problem, possibly caused by not considering theory-of-mind theory (ToM). Inspired by the ToM personality in cognitive psychology, where a human can easily resolve this problem by predicting others' intuitive personality first before complex actions, we propose a biologically-plausible algorithm named the mixture of personality (MoP) improved spiking actor network (SAN). The MoP module contains a determinantal point process to simulate the formation and integration of different personality types, and the SAN module contains spiking neurons for efficient reinforcement learning. The experimental results on the benchmark cooperative overcooked task showed that the proposed MoP-SAN algorithm could achieve higher performance for the paradigms with (learning) and without (generalization) unseen partners. Furthermore, ablation experiments highlighted the contribution of MoP in SAN learning, and some visualization analysis explained why the proposed algorithm is superior to some counterpart deep actor networks.

Funder

Youth Innovation Promotion Association of the Chinese Academy of Sciences

Publisher

Frontiers Media SA

Subject

General Neuroscience

Reference51 articles.

1. Effect of the COVID-19 pandemic and big five personality on subjective and psychological well-being;Anglim;Soc. Psychol. Pers. Sci.,2021

2. Mind the gap: challenges of deep learning approaches to theory of mind;Aru;Artif. Intell. Rev.,2023

3. A solution to the learning dilemma for recurrent networks of spiking neurons;Bellec;Nat. Commun.,2020

4. Culture and the evolution of human cooperation;Boyd;Philos. Trans. R. Soc. B Biol. Sci.,2009

5. “On the utility of learning about humans for Human-AI coordination,” CarrollM. ShahR. HoM. K. GriffithsT. SeshiaS. AbbeelP. Advances in Neural Information Processing Systems2019

Cited by 4 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. S²RC-GCN: A Spatial-Spectral Reliable Contrastive Graph Convolutional Network for Complex Land Cover Classification Using Hyperspectral Images;2024 International Joint Conference on Neural Networks (IJCNN);2024-06-30

2. Shared and Private Information Learning in Multimodal Sentiment Analysis with Deep Modal Alignment and Self-supervised Multi-Task Learning;2024 International Joint Conference on Neural Networks (IJCNN);2024-06-30

3. Multi-level Graph Subspace Contrastive Learning for Hyperspectral Image Clustering;2024 International Joint Conference on Neural Networks (IJCNN);2024-06-30

4. Long Short-Term Reasoning Network with Theory of Mind for Efficient Multi-Agent Cooperation;2024 International Joint Conference on Neural Networks (IJCNN);2024-06-30