Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Seyed Sajad Mousavi

Deep Reinforcement Learning: An Overview

Jun 23, 2018

Seyed Sajad Mousavi, Michael Schukat, Enda Howley

Figure 1 for Deep Reinforcement Learning: An Overview

Abstract:In recent years, a specific machine learning method called deep learning has gained huge attraction, as it has obtained astonishing results in broad applications such as pattern recognition, speech recognition, computer vision, and natural language processing. Recent research has also been shown that deep learning techniques can be combined with reinforcement learning methods to learn useful representations for the problems with high dimensional raw data input. This chapter reviews the recent advances in deep reinforcement learning with a focus on the most used deep architectures such as autoencoders, convolutional neural networks and recurrent neural networks which have successfully been come together with the reinforcement learning framework.

* Proceedings of SAI Intelligent Systems Conference (IntelliSys) 2016

Via

Access Paper or Ask Questions

Traffic Light Control Using Deep Policy-Gradient and Value-Function Based Reinforcement Learning

May 27, 2017

Seyed Sajad Mousavi, Michael Schukat, Enda Howley

Figure 1 for Traffic Light Control Using Deep Policy-Gradient and Value-Function Based Reinforcement Learning

Figure 2 for Traffic Light Control Using Deep Policy-Gradient and Value-Function Based Reinforcement Learning

Figure 3 for Traffic Light Control Using Deep Policy-Gradient and Value-Function Based Reinforcement Learning

Figure 4 for Traffic Light Control Using Deep Policy-Gradient and Value-Function Based Reinforcement Learning

Abstract:Recent advances in combining deep neural network architectures with reinforcement learning techniques have shown promising potential results in solving complex control problems with high dimensional state and action spaces. Inspired by these successes, in this paper, we build two kinds of reinforcement learning algorithms: deep policy-gradient and value-function based agents which can predict the best possible traffic signal for a traffic intersection. At each time step, these adaptive traffic light control agents receive a snapshot of the current state of a graphical traffic simulator and produce control signals. The policy-gradient based agent maps its observation directly to the control signal, however the value-function based agent first estimates values for all legal control signals. The agent then selects the optimal control action with the highest value. Our methods show promising results in a traffic network simulated in the SUMO traffic simulator, without suffering from instability issues during the training process.

Via

Access Paper or Ask Questions