Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Sample-Efficient Deep RL with Generative Adversarial Tree Search

Jun 15, 2018

Kamyar Azizzadenesheli, Brandon Yang, Weitang Liu, Emma Brunskill, Zachary C Lipton, Animashree Anandkumar

Figure 1 for Sample-Efficient Deep RL with Generative Adversarial Tree Search

Figure 2 for Sample-Efficient Deep RL with Generative Adversarial Tree Search

Figure 3 for Sample-Efficient Deep RL with Generative Adversarial Tree Search

Figure 4 for Sample-Efficient Deep RL with Generative Adversarial Tree Search

Share this with someone who'll enjoy it:

Abstract:We propose Generative Adversarial Tree Search (GATS), a sample-efficient Deep Reinforcement Learning (DRL) algorithm. While Monte Carlo Tree Search (MCTS) is known to be effective for search and planning in RL, it is often sample-inefficient and therefore expensive to apply in practice. In this work, we develop a Generative Adversarial Network (GAN) architecture to model an environment's dynamics and a predictor model for the reward function. We exploit collected data from interaction with the environment to learn these models, which we then use for model-based planning. During planning, we deploy a finite depth MCTS, using the learned model for tree search and a learned Q-value for the leaves, to find the best action. We theoretically show that GATS improves the bias-variance trade-off in value-based DRL. Moreover, we show that the generative model learns the model dynamics using orders of magnitude fewer samples than the Q-learner. In non-stationary settings where the environment model changes, we find the generative model adapts significantly faster than the Q-learner to the new environment.

View paper on

Share this with someone who'll enjoy it:

Title:Sample-Efficient Deep RL with Generative Adversarial Tree Search

Paper and Code