Making Teams and Influencing Agents: Efficiently Coordinating Decision Trees for Interpretable Multi-Agent Reinforcement Learning

Authors

  • Rex Chen School of Computer Science, Carnegie Mellon University
  • Stephanie Milani School of Computer Science, Carnegie Mellon University
  • Zhicheng Zhang School of Computer Science, Carnegie Mellon University
  • Norman Sadeh School of Computer Science, Carnegie Mellon University
  • Fei Fang School of Computer Science, Carnegie Mellon University

DOI:

https://doi.org/10.1609/aies.v8i1.36571

Abstract

Poor interpretability hinders the practical applicability of multi-agent reinforcement learning (MARL) policies. Deploying interpretable surrogates of uninterpretable policies enhances the safety and verifiability of MARL for real-world applications. However, if these surrogates are to interact directly with the environment within human supervisory frameworks, they must be both performant and computationally efficient. Prior work on interpretable MARL has either sacrificed performance for computational efficiency or computational efficiency for performance. To address this issue, we propose HYDRAVIPER, a decision tree-based interpretable MARL algorithm. HYDRAVIPER coordinates training between agents based on expected team performance, and adaptively allocates budgets for environment interaction to improve computational efficiency. Experiments on standard benchmark environments for multi-agent coordination and traffic signal control show that HYDRAVIPER matches the performance of state-of-the-art methods using a fraction of the runtime, and that it maintains a Pareto frontier of performance for different interaction budgets.

Downloads

Published

2025-10-15

How to Cite

Chen, R., Milani, S., Zhang, Z., Sadeh, N., & Fang, F. (2025). Making Teams and Influencing Agents: Efficiently Coordinating Decision Trees for Interpretable Multi-Agent Reinforcement Learning. Proceedings of the AAAI ACM Conference on AI, Ethics, and Society, 8(1), 567–578. https://doi.org/10.1609/aies.v8i1.36571