Computer Science > Machine Learning

arXiv:2205.12449 (cs)

[Submitted on 25 May 2022 (v1), last revised 11 Jul 2022 (this version, v2)]

Title:MAVIPER: Learning Decision Tree Policies for Interpretable Multi-Agent Reinforcement Learning

Authors:Stephanie Milani, Zhicheng Zhang, Nicholay Topin, Zheyuan Ryan Shi, Charles Kamhoua, Evangelos E. Papalexakis, Fei Fang

View PDF

Abstract:Many recent breakthroughs in multi-agent reinforcement learning (MARL) require the use of deep neural networks, which are challenging for human experts to interpret and understand. On the other hand, existing work on interpretable reinforcement learning (RL) has shown promise in extracting more interpretable decision tree-based policies from neural networks, but only in the single-agent setting. To fill this gap, we propose the first set of algorithms that extract interpretable decision-tree policies from neural networks trained with MARL. The first algorithm, IVIPER, extends VIPER, a recent method for single-agent interpretable RL, to the multi-agent setting. We demonstrate that IVIPER learns high-quality decision-tree policies for each agent. To better capture coordination between agents, we propose a novel centralized decision-tree training algorithm, MAVIPER. MAVIPER jointly grows the trees of each agent by predicting the behavior of the other agents using their anticipated trees, and uses resampling to focus on states that are critical for its interactions with other agents. We show that both algorithms generally outperform the baselines and that MAVIPER-trained agents achieve better-coordinated performance than IVIPER-trained agents on three different multi-agent particle-world environments.

Comments:	ECML camera-ready version. 23 pages
Subjects:	Machine Learning (cs.LG); Multiagent Systems (cs.MA)
Cite as:	arXiv:2205.12449 [cs.LG]
	(or arXiv:2205.12449v2 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2205.12449

Submission history

From: Stephanie Milani [view email]
[v1] Wed, 25 May 2022 02:38:10 UTC (1,234 KB)
[v2] Mon, 11 Jul 2022 20:46:36 UTC (9,233 KB)

Computer Science > Machine Learning

Title:MAVIPER: Learning Decision Tree Policies for Interpretable Multi-Agent Reinforcement Learning

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:MAVIPER: Learning Decision Tree Policies for Interpretable Multi-Agent Reinforcement Learning

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators