Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Claude Klöckl

RangL: A Reinforcement Learning Competition Platform

Jul 28, 2022

Viktor Zobernig, Richard A. Saldanha, Jinke He, Erica van der Sar, Jasper van Doorn, Jia-Chen Hua, Lachlan R. Mason, Aleksander Czechowski, Drago Indjic, Tomasz Kosmala(+5 more)

Figure 1 for RangL: A Reinforcement Learning Competition Platform

Figure 2 for RangL: A Reinforcement Learning Competition Platform

Figure 3 for RangL: A Reinforcement Learning Competition Platform

Abstract:The RangL project hosted by The Alan Turing Institute aims to encourage the wider uptake of reinforcement learning by supporting competitions relating to real-world dynamic decision problems. This article describes the reusable code repository developed by the RangL team and deployed for the 2022 Pathways to Net Zero Challenge, supported by the UK Net Zero Technology Centre. The winning solutions to this particular Challenge seek to optimize the UK's energy transition policy to net zero carbon emissions by 2050. The RangL repository includes an OpenAI Gym reinforcement learning environment and code that supports both submission to, and evaluation in, a remote instance of the open source EvalAI platform as well as all winning learning agent strategies. The repository is an illustrative example of RangL's capability to provide a reusable structure for future challenges.

* Documents in general and premierly the RangL competition plattform and in particular its 2022's competition "Pathways to Netzero" 10 pages, 2 figures, 1 table, Comments welcome!

Via

Access Paper or Ask Questions

Computational Performance of Deep Reinforcement Learning to find Nash Equilibria

Apr 26, 2021

Christoph Graf, Viktor Zobernig, Johannes Schmidt, Claude Klöckl

Figure 1 for Computational Performance of Deep Reinforcement Learning to find Nash Equilibria

Figure 2 for Computational Performance of Deep Reinforcement Learning to find Nash Equilibria

Figure 3 for Computational Performance of Deep Reinforcement Learning to find Nash Equilibria

Figure 4 for Computational Performance of Deep Reinforcement Learning to find Nash Equilibria

Abstract:We test the performance of deep deterministic policy gradient (DDPG), a deep reinforcement learning algorithm, able to handle continuous state and action spaces, to learn Nash equilibria in a setting where firms compete in prices. These algorithms are typically considered model-free because they do not require transition probability functions (as in e.g., Markov games) or predefined functional forms. Despite being model-free, a large set of parameters are utilized in various steps of the algorithm. These are e.g., learning rates, memory buffers, state-space dimensioning, normalizations, or noise decay rates and the purpose of this work is to systematically test the effect of these parameter configurations on convergence to the analytically derived Bertrand equilibrium. We find parameter choices that can reach convergence rates of up to 99%. The reliable convergence may make the method a useful tool to study strategic behavior of firms even in more complex settings. Keywords: Bertrand Equilibrium, Competition in Uniform Price Auctions, Deep Deterministic Policy Gradient Algorithm, Parameter Sensitivity Analysis

* 48 pages + 9 figures, comments welcome!

Via

Access Paper or Ask Questions