The StarCraft Multi-Agent Challenge

2019-02-11Code Available1· sign in to hype

Mikayel Samvelyan, Tabish Rashid, Christian Schroeder de Witt, Gregory Farquhar, Nantas Nardelli, Tim G. J. Rudner, Chia-Man Hung, Philip H. S. Torr, Jakob Foerster, Shimon Whiteson

arXiv PDF

Code Available — Be the first to reproduce this paper.

Reproduce

Code

github.com/oxwhirl/pymarl
OfficialIn paperpytorch★ 2,168
github.com/oxwhirl/smac
OfficialIn paperpytorch★ 0
github.com/uoe-agents/epymarl
pytorch★ 702
github.com/oxwhirl/smacv2
none★ 299
github.com/oxwhirl/facmac
pytorch★ 111
github.com/ailabdsunipi/pymarlzooplus
pytorch★ 47
github.com/hahayonghuming/VDACs
pytorch★ 42
github.com/osilab-kaist/smac_exp
pytorch★ 31
github.com/ling-pan/res
pytorch★ 25
github.com/jk96491/C-COMA
pytorch★ 11

Abstract

In the last few years, deep multi-agent reinforcement learning (RL) has become a highly active area of research. A particularly challenging class of problems in this area is partially observable, cooperative, multi-agent learning, in which teams of agents must learn to coordinate their behaviour while conditioning only on their private observations. This is an attractive research area since such problems are relevant to a large number of real-world systems and are also more amenable to evaluation than general-sum problems. Standardised environments such as the ALE and MuJoCo have allowed single-agent RL to move beyond toy domains, such as grid worlds. However, there is no comparable benchmark for cooperative multi-agent RL. As a result, most papers in this field use one-off toy problems, making it difficult to measure real progress. In this paper, we propose the StarCraft Multi-Agent Challenge (SMAC) as a benchmark problem to fill this gap. SMAC is based on the popular real-time strategy game StarCraft II and focuses on micromanagement challenges where each unit is controlled by an independent agent that must act based on local observations. We offer a diverse set of challenge maps and recommendations for best practices in benchmarking and evaluations. We also open-source a deep multi-agent RL learning framework including state-of-the-art algorithms. We believe that SMAC can provide a standard benchmark environment for years to come. Videos of our best agents for several SMAC scenarios are available at: https://youtu.be/VZ7zmQ_obZ0.

Tasks

Benchmarking MuJoCo Multi-agent Reinforcement Learning Real-Time Strategy Games Reinforcement Learning Reinforcement Learning (RL)SMAC SMAC+Starcraft Starcraft II

Benchmark Results

Dataset	Model	Metric	Claimed	Verified	Status
SMAC 27m_vs_30m	Heuristic	Median Win Rate	0	—	Unverified
SMAC 3s5z_vs_3s6z	VDN	Median Win Rate	2	—	Unverified
SMAC 3s5z_vs_3s6z	IQL	Median Win Rate	0	—	Unverified
SMAC 3s5z_vs_3s6z	VDN	Median Win Rate	89.2	—	Unverified
SMAC 3s5z_vs_3s6z	IQL	Median Win Rate	29.83	—	Unverified
SMAC 3s5z_vs_3s6z	Heuristic	Median Win Rate	0	—	Unverified
SMAC 6h_vs_8z	Heuristic	Median Win Rate	0	—	Unverified
SMAC 6h_vs_8z	VDN	Median Win Rate	0	—	Unverified
SMAC 6h_vs_8z	IQL	Median Win Rate	0	—	Unverified
SMAC corridor	Heuristic	Median Win Rate	0	—	Unverified
SMAC corridor	IQL	Median Win Rate	0	—	Unverified
SMAC corridor	IQL	Median Win Rate	84.87	—	Unverified
SMAC MMM2	Heuristic	Median Win Rate	0	—	Unverified
SMAC MMM2	VDN	Median Win Rate	1	—	Unverified
SMAC MMM2	IQL	Median Win Rate	0	—	Unverified
SMAC MMM2	VDN	Median Win Rate	89.2	—	Unverified
SMAC MMM2	IQL	Median Win Rate	68.92	—	Unverified

The StarCraft Multi-Agent Challenge

Code

Abstract

Tasks

Benchmark Results

Reproductions