Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Learning to Play Atari in a World of Tokens

Jun 03, 2024

Pranav Agarwal, Sheldon Andrews, Samira Ebrahimi Kahou

Figure 1 for Learning to Play Atari in a World of Tokens

Figure 2 for Learning to Play Atari in a World of Tokens

Figure 3 for Learning to Play Atari in a World of Tokens

Figure 4 for Learning to Play Atari in a World of Tokens

Share this with someone who'll enjoy it:

Abstract:Model-based reinforcement learning agents utilizing transformers have shown improved sample efficiency due to their ability to model extended context, resulting in more accurate world models. However, for complex reasoning and planning tasks, these methods primarily rely on continuous representations. This complicates modeling of discrete properties of the real world such as disjoint object classes between which interpolation is not plausible. In this work, we introduce discrete abstract representations for transformer-based learning (DART), a sample-efficient method utilizing discrete representations for modeling both the world and learning behavior. We incorporate a transformer-decoder for auto-regressive world modeling and a transformer-encoder for learning behavior by attending to task-relevant cues in the discrete representation of the world model. For handling partial observability, we aggregate information from past time steps as memory tokens. DART outperforms previous state-of-the-art methods that do not use look-ahead search on the Atari 100k sample efficiency benchmark with a median human-normalized score of 0.790 and beats humans in 9 out of 26 games. We release our code at https://pranaval.github.io/DART/.

* Accepted at ICML 2024

View paper on

Share this with someone who'll enjoy it:

Title:Learning to Play Atari in a World of Tokens

Paper and Code