Full Communication Memory Networks for Team-Level Cooperation Learning

doi:10.21203/rs.3.rs-2563058/v1

Download PDF

Research Article

Full Communication Memory Networks for Team-Level Cooperation Learning

https://doi.org/10.21203/rs.3.rs-2563058/v1

This work is licensed under a CC BY 4.0 License

Journal Publication

published 07 Aug, 2023

Read the published version in Autonomous Agents and Multi-Agent Systems →

You are reading this latest preprint version

Communication in multi-agent systems is a key driver of team-level cooperation, for instance allowing individual agents to augment their knowledge about the world in partially-observable environments. In this paper, we propose two reinforcement learning-based multi-agent models, namely FCMNet and FCMTran. The two models both allow agents to simultaneously learn a differentiable communication mechanism that connects all agents as well as a common, cooperative policy conditioned upon received information. FCMNet utilizes multiple directional recurrent neural networks to sequentially transmit and encode the current observation-based messages sent by every other agent at each timestep. FCMTran further relies on the encoder of a modified transformer to simultaneously aggregate multiple self-generated messages sent by all agents at the previous timestep into a single message that is used in the current timestep. Results from evaluating our models on a challenging set of StarCraft II micromanagement tasks with shared rewards show that FCMNet and FCMTran both outperform recent communication-based methods and value decomposition methods in almost all tested StarCraft II micromanagement tasks. We further improve the performance of our models by combining them with value decomposition techniques; there, in particular, we show that FCMTran with value decomposition significantly pushes the state-of-the-art on one of the hardest benchmark tasks without any task-specific tuning. We also investigate the robustness of FCMNet under communication disturbances (i.e., binarized messages, random message loss, and random communication order) in an asymmetric collaborative pathfinding task with individual rewards, demonstrating FMCNet’s potential applicability in real-world robotic tasks.

Multi-Agent System

Reinforcement Learning

Differentiable Communications

Decentralized Cooperation

No competing interests reported.

Supplementarymaterial.zip

Download PDF

Journal Publication

published 07 Aug, 2023

Read the published version in Autonomous Agents and Multi-Agent Systems →

Editorial decision: Major revision
20 May, 2023
Reviews received at journal
25 Apr, 2023
Reviewers agreed at journal
18 Apr, 2023
Reviewers agreed at journal
20 Feb, 2023
Reviewers invited by journal
17 Feb, 2023
Editor assigned by journal
08 Feb, 2023
Submission checks completed at journal
08 Feb, 2023
First submitted to journal
08 Feb, 2023

You are reading this latest preprint version

Full Communication Memory Networks for Team-Level Cooperation Learning

Status:

Journal Publication

Version 1

Abstract

Full Text

Additional Declarations

Supplementary Files

Status:

Journal Publication

Version 1