Multiagent Bidirectionally-Coordinated Nets: Emergence of Human-level   Coordination in Learning to Play StarCraft Combat Games

Peng Peng; Ying Wen; Yaodong Yang; Quan Yuan; Zhenkun Tang; Haitao; Long; Jun Wang

arXiv:1703.10069·cs.AI·September 15, 2017·272 cites

Multiagent Bidirectionally-Coordinated Nets: Emergence of Human-level Coordination in Learning to Play StarCraft Combat Games

Peng Peng, Ying Wen, Yaodong Yang, Quan Yuan, Zhenkun Tang, Haitao, Long, Jun Wang

PDF

Open Access 2 Repos

TL;DR

This paper introduces BiCNet, a multiagent neural network that learns human-level coordination strategies in StarCraft combat games without supervision, demonstrating state-of-the-art performance and scalability.

Contribution

The paper presents BiCNet, a novel bidirectional communication network enabling scalable, unsupervised learning of complex multiagent coordination in real-time strategy games.

Findings

01

BiCNet achieves state-of-the-art performance in StarCraft combat scenarios.

02

It learns advanced coordination strategies without supervision.

03

The approach scales to arbitrary numbers of agents.

Abstract

Many artificial intelligence (AI) applications often require multiple intelligent agents to work in a collaborative effort. Efficient learning for intra-agent communication and coordination is an indispensable step towards general AI. In this paper, we take StarCraft combat game as a case study, where the task is to coordinate multiple agents as a team to defeat their enemies. To maintain a scalable yet effective communication protocol, we introduce a Multiagent Bidirectionally-Coordinated Network (BiCNet ['bIknet]) with a vectorised extension of actor-critic formulation. We show that BiCNet can handle different types of combats with arbitrary numbers of AI agents for both sides. Our analysis demonstrates that without any supervisions such as human demonstrations or labelled data, BiCNet could learn various types of advanced coordination strategies that have been commonly used by…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsReinforcement Learning in Robotics · Adversarial Robustness in Machine Learning · Artificial Intelligence in Games