Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Michael Walton

Multi-agent Reinforcement Learning in OpenSpiel: A Reproduction Report

Mar 02, 2021

Michael Walton, Viliam Lisy

Figure 1 for Multi-agent Reinforcement Learning in OpenSpiel: A Reproduction Report

Figure 2 for Multi-agent Reinforcement Learning in OpenSpiel: A Reproduction Report

Figure 3 for Multi-agent Reinforcement Learning in OpenSpiel: A Reproduction Report

Figure 4 for Multi-agent Reinforcement Learning in OpenSpiel: A Reproduction Report

Abstract:In this report, we present results reproductions for several core algorithms implemented in the OpenSpiel framework for learning in games. The primary contribution of this work is a validation of OpenSpiel's re-implemented search and Reinforcement Learning algorithms against the results reported in their respective originating works. Additionally, we provide complete documentation of hyperparameters and source code required to reproduce these experiments easily and exactly.

Via

Access Paper or Ask Questions

Theory of Mind for Deep Reinforcement Learning in Hanabi

Jan 22, 2021

Andrew Fuchs, Michael Walton, Theresa Chadwick, Doug Lange

Figure 1 for Theory of Mind for Deep Reinforcement Learning in Hanabi

Figure 2 for Theory of Mind for Deep Reinforcement Learning in Hanabi

Figure 3 for Theory of Mind for Deep Reinforcement Learning in Hanabi

Figure 4 for Theory of Mind for Deep Reinforcement Learning in Hanabi

Abstract:The partially observable card game Hanabi has recently been proposed as a new AI challenge problem due to its dependence on implicit communication conventions and apparent necessity of theory of mind reasoning for efficient play. In this work, we propose a mechanism for imbuing Reinforcement Learning agents with a theory of mind to discover efficient cooperative strategies in Hanabi. The primary contributions of this work are threefold: First, a formal definition of a computationally tractable mechanism for computing hand probabilities in Hanabi. Second, an extension to conventional Deep Reinforcement Learning that introduces reasoning over finitely nested theory of mind belief hierarchies. Finally, an intrinsic reward mechanism enabled by theory of mind that incentivizes agents to share strategically relevant private knowledge with their teammates. We demonstrate the utility of our algorithm against Rainbow, a state-of-the-art Reinforcement Learning agent.

Via

Access Paper or Ask Questions