Efficient Communication via Self-supervised Information Aggregation for Online and Offline Multi-agent Reinforcement Learning

02/19/2023
by   Cong Guan, et al.
0

Utilizing messages from teammates can improve coordination in cooperative Multi-agent Reinforcement Learning (MARL). Previous works typically combine raw messages of teammates with local information as inputs for policy. However, neglecting message aggregation poses significant inefficiency for policy learning. Motivated by recent advances in representation learning, we argue that efficient message aggregation is essential for good coordination in cooperative MARL. In this paper, we propose Multi-Agent communication via Self-supervised Information Aggregation (MASIA), where agents can aggregate the received messages into compact representations with high relevance to augment the local policy. Specifically, we design a permutation invariant message encoder to generate common information-aggregated representation from messages and optimize it via reconstructing and shooting future information in a self-supervised manner. Hence, each agent would utilize the most relevant parts of the aggregated representation for decision-making by a novel message extraction mechanism. Furthermore, considering the potential of offline learning for real-world applications, we build offline benchmarks for multi-agent communication, which is the first as we know. Empirical results demonstrate the superiority of our method in both online and offline settings. We also release the built offline benchmarks in this paper as a testbed for communication ability validation to facilitate further future research.

READ FULL TEXT

page 7

page 8

page 9

page 15

page 28

page 29

page 30

page 31

research
05/23/2023

Research on Multi-Agent Communication and Collaborative Decision-Making Based on Deep Reinforcement Learning

In a multi-agent environment, In order to overcome and alleviate the non...
research
08/10/2022

Diversifying Message Aggregation in Multi-Agent Communication via Normalized Tensor Nuclear Norm Regularization

Aggregating messages is a key component for the communication of multi-a...
research
06/16/2023

Dynamic Size Message Scheduling for Multi-Agent Communication under Limited Bandwidth

Communication plays a vital role in multi-agent systems, fostering colla...
research
05/07/2023

Robust Multi-agent Communication via Multi-view Message Certification

Many multi-agent scenarios require message sharing among agents to promo...
research
12/03/2019

Learning Agent Communication under Limited Bandwidth by Message Pruning

Communication is a crucial factor for the big multi-agent world to stay ...
research
10/02/2020

Correcting Experience Replay for Multi-Agent Communication

We consider the problem of learning to communicate using multi-agent rei...
research
03/19/2023

Cheap Talk Discovery and Utilization in Multi-Agent Reinforcement Learning

By enabling agents to communicate, recent cooperative multi-agent reinfo...

Please sign up or login with your details

Forgot password? Click here to reset